Alexandra Instituttet NLP & DaCY
Production-ready Danish language pipelines, named entity recognition (DaNE), and benchmarking suites (ScandEval).
About
Alexandra Instituttet develops and maintains core open-source NLP toolkits and benchmark infrastructure for the Danish language. Key assets include DaCY (a state-of-the-art Danish NLP framework built on spaCy), DaNE (the standard Danish Named Entity dataset covering persons, locations, and organizations), and ScandEval (the standard Scandinavian LLM evaluation leaderboard).
Actionable Use Cases
- 01Extract Danish named entities (personer, organisationer, lokationer) from unstructured text
- 02Benchmark and evaluate LLM performance on standardized Danish comprehension tasks via ScandEval
- 03Tokenize, POS-tag, and parse Danish sentence dependencies in Python
Technical Details & Limits
Fully open-source Python packages distributed via PyPI and Hugging Face. Excellent documentation and research reproducibility.
Build Access with Resources
Turn this API into an MCP server or CLI for Claude, Cursor, and terminal agents:
Instant MCP & CLI generator from OpenAPI, Swagger & APIs
High-level Python & TypeScript framework for building MCP servers
Turn entire websites and portals into clean, LLM-ready Markdown
Something wrong, outdated, or missing for Alexandra Instituttet NLP & DaCY?