Skip to content

open-models

Empero releases Qwythos-9B trained on Claude traces

Claude News

Laboratory Empero has introduced Qwythos-9B, a reasoning model based on Qwen3.5-9B, fine-tuned on over 500 million tokens of Claude Mythos and Claude Fable traces. The model uses full-parameter fine-tuning and is released under an Apache 2.0 license.

Using the lm-evaluation-harness, the model outperforms the base Qwen3.5-9B: +34.3 points on MMLU (0.575 vs 0.232), +30 points on gsm8k-strict, and +19 on gsm8k-flex. On gpqa_diamond, the result is 5 points lower than the base model.

The context window is extended to 1,048,576 tokens via YaRN rope-scaling with a factor of 4, built on top of the native 262,144 tokens. It includes native function calling: in a 7-prompt test involving a Python executor and web search, the model provided correct answers with links in every case.

The model is intentionally uncensored and optimized for cybersecurity, biomedicine, and pharmacology. Weights are available on Hugging Face.

Related stories

  1. Claude Opus 5.5 cut PSP math tables from 4.9 MB to 10.5 KB
  2. Opus 5.5 leads the index and burns 260M tokens doing it
  3. Opus 5.5 takes the default seat in Claude Code
  4. Opus 5.5 matches Fable 5.1 on most work for less money
  5. Max20x buys only 1.5x of Max5x, says one team's tally
  6. Claude Fable 5.2 draws its sprites in raw JavaScript

Comments

No comments yet. Be the first.

Join the conversation

Sign in with Google to leave a comment. Your name and avatar come from your Google profile, and the comment appears after moderation.

We only use your name and avatar from Google. We never store your email address.