Lukas Edman
leukas
AI & ML interests
NLP, tokenization, pretraining, low-resource
Recent Activity
updated a model 1 day ago
leukas/Qwenfant-0.4B updated a dataset 1 day ago
leukas/Qwenfood updated a model 1 day ago
leukas/Qwenfant-0.1BOrganizations
Are Character-level Translations Worth the Wait?
Collection of trained models for the paper: Are Character-level Translations Worth the Wait?
BabyLM 2023
Final model for BabyLM 2023
BabyLM 2026
BabyLM 2025
Are Character-level Translations Worth the Wait?
Collection of trained models for the paper: Are Character-level Translations Worth the Wait?
BabyLM 2024
BabyLM2024 models
BabyLM 2023
Final model for BabyLM 2023
CUTE
The CUTE benchmark is an LLM benchmark, testing LLMs' understanding of orthography. Check out our github here: https://github.com/Leukas/cute