
OpenEuroLLM at Cyber Valley Day
Cyber Valley Innovation Campus
Our colleagues from ELLIS Institute Tübingen and Tübingen AI Center were presenting the OpenEuroLLM project at the Cyber Valley Innovation Campus 10th anniversary at Tübingen
including data, documentation, training and testing code, and evaluation metrics; including community involvement
under EU regulations, OpenEuroLLM will provide a series of transparent and performant LLMs
for European languages and other socially and economically interesting ones, preserving linguistic and cultural diversity
Updates
Develop strong multilingual foundation models for EU official languages and beyond.
Ensure easy and sustainable access to foundational models ready to be fine-tuned to a wide range of applications.
Extend the number of evaluation results in all EU official languages and beyond, including AI safety and alignment with the AI Act and European AI standards
Extend the number of available training datasets and benchmarks for these languages, and make them easily accessible.
Share transparently the tools, recipes and intermediate results of the training processes.
Share the dataset enrichment and anonymization pipelines to enable further data sourcing for future needs.
Create an active and engaged community of developers and stakeholders among the public and private sector.
Find answers about OpenEuroLLM code, data, models, releases, research progress and multilingual coverage.
The code we wrote is available in https://github.com/OpenEuroLLM In particular,
Europe's leading AI companies and research institutions combine their forces and expertise to develop next-generation open-source language models in an unprecedented collaboration to advance European AI capabilities, the OpenEuroLLM project