There is a child with a rare disease who is currently suffering and struggling to manage his symptoms. Rare as this is, you can directly help him.
Today we are launching the "Rare Disease, Real Kid" Hackathon, and there are $50,000 in prizes from @AnthropicAI and @awscloud.
We
Sol> result is A
me> A is wrong because of [reason]
Sol> Correct. A is wrong because of [reason rephrased]. You should do B instead.
me> B is also wrong because of [reason]
Sol> Agreed. You should not do B because of [reason rephrased].
Ox Alpha achieves 26.6% on TerminalBench 3.0!
This puts it above Opus 4.8 but behind Fable 5, between GLM 5.3 and Grok 4.6. Based on the fingerprints and results, makes a lot of sense for this to be GLM 5.x Flash with vision capability.
Caveat: This is 1x pass@1, whereas the
🥷 New stealth model: Ox Alpha
Ox Alpha is a frontier model built for efficient coding, sustained agentic work, and real-world production use.
- 1M token context window
- Text, image, and video input
Try it now and share feedback to improve the model! openrouter.ai/stealth/ox-alp…
Best Qwen3.8 27B serving config found for RTX PRO 6000!
SGLang / RadixArk NVFP4 / Inco DFlash 2 / FP8 KV
I was only getting ~130 tok/s on coding workload before even with dspark/dflash 2. With this config I am hitting 170+ on non-toy benchmarks.
@sgl_project@Alibaba_Qwen
Just pushed DFlash2 (@inco_ai) recipes to the Qwen3.8 27B cookbook⚡️
docs.sglang.io/cookbook/autor…
The community has been seeing great results with NVFP4 + DFlash2, and these recipes should be some very good starting points to play with.
More Qwen3.8 27B updates on the way 🫡