Open Source Projects
Turkish Morphological Tokenizer
A modern tokenizer that splits text into morphological units faithful to Turkish phonetics and can recombine them.
View project →Language-Native Embeddings
An open methodology compiling methods and steps for anyone to build efficient tokenizers and embedding models for their own language and domain.
View project →Fine-tune Datasets
A community project gathering high-quality Turkish/English dialogue datasets in a single format to adapt models to a task, language, or persona.
View project →Magibu LLM Tools
Tools and auxiliary systems the model uses at inference time - overcoming model limits by calling the right tool at the right time.
View project →Community
Open-Ended Live Stream
A Sunday live stream with no end time: code, questions, papers. Leave a request for the topics you want covered.
Request form →Magibu AI Weekly
Open-source weekly digest: AI news, papers, models, benchmarks, and underrepresented language updates.
View archive →