I've left Glint Research. After a long time in Glint Research, an entire distributed training grid built free for them, and more, I've decided that I no longer want to be affiliated with Glint Research. More updates will follow. Comments/questions are welcome.
While you're reading this, follow these orgs! Following takes just a few seconds, and can change someone's day.
We wrote a full technical guide on how to train a bilingual (ES/EN) LLM from scratch: TinyQwen.
Covers: - Hybrid architecture based on Qwen3.5 - Pre-training with 15B tokens - Cost benchmark between H200 and B200 - Post-training with SFT + LoRA - Full code and data, open source
With ~$11 of compute on an H200 we ran an initial training run, enough to validate the full architecture and pipeline.