What is K2 Horizon 0.9B?
K2 Horizon 0.9B is IFM's smallest dense reasoning model. It accepts text and uses multi-teacher distillation to combine math, coding and instruction-following skills in a compact checkpoint. You can evaluate it for short code generation or tightly scoped tool selection before moving to a larger model.
| Specification | K2 Horizon 0.9B |
|---|---|
| Parameters | 0.9B class; 1.078B stored parameters |
| Architecture | Dense decoder-only |
| Layers | 28 |
| Context window | 131,072 tokens; YaRN from 8,192 |
| Modalities | Text input and output |
| Release status | Released checkpoint |
| License | Apache-2.0 label; conflicting metadata |
The 128K window comes from YaRN scaling beyond the original 8,192-token context. It is not native 128K pretraining. The 0.9B name is also a size class: Hugging Face reports 1.078B stored parameters. Use the actual file inventory when choosing a download, especially on a memory-constrained machine.
K2 Horizon 0.9B benchmarks
IFM reports these percentage scores in the model card. Qwen3.5-2B is a larger reference model. They are vendor-reported results, not tests run in Atomic Chat; the card links a technical appendix for protocol details. Source: official model card.
| Benchmark | K2-Horizon-0.9B | Qwen3.5-0.8B | OpenBMB-1B | Qwen3.5-2B |
|---|---|---|---|---|
AIME 2025 Competition math | 41.7 | 1.0 | 40.4 | 34.2 |
AIME 2026 Competition math | 48.5 | 0.2 | 40.4 | 38.8 |
HMMT Feb 2026 Olympiad math | 25.8 | 0.6 | 23.3 | 22.7 |
GPQA Diamond Expert science | 27.3 | 11.9 | 26.3 | 54.9 |
HumanEval+ Code generation | 79.9 | 16.5 | 65.2 | 75.6 |
BFCL v4 Function calling | 28.0 | 25.3 | 25.2 | 43.6 |
The 0.9B model leads these references on the selected math tests and HumanEval+. Qwen3.5-2B is ahead on GPQA Diamond and BFCL v4. If your application depends on reliable function selection, the coding scores alone are not enough to choose the smaller model.
K2 Horizon 0.9B hardware requirements
The original Safetensors download is 2.16 GB. IFM also publishes a BF16 GGUF whose filename says 1B, while its repository identifies the source as K2 Horizon 0.9B. This is a format conversion, not a low-bit quantization.
| Precision | Source | Size on disk |
|---|---|---|
| BF16 Safetensors | IFM/K2-Horizon-0.9B | 2.16 GB |
| BF16 GGUF (official) | IFM/K2-Horizon-0.9B-GGUF K2-Horizon-1B-BF16.gguf | 2.16 GB |
A 2 GB memory pool cannot hold the complete BF16 weight set plus the inference engine. The GGUF file size does not include the operating system or a growing KV cache. Start capacity testing at a short context; the advertised 128K maximum is not a measured fit on a small device.
Sizes use decimal GB and cover weights only. For the format distinction, see what GGUF is. A memory tier is not listed because minimum runtime memory has not been measured for these builds.
How to run K2 Horizon 0.9B in Atomic Chat
Atomic Chat execution has not been verified for this checkpoint. IFM's official GGUF card requires a llama.cpp build with K2 Horizon architecture support and points to its development fork. Check architecture support before downloading these files. The usual search, download and start-chat flow is not yet a verified procedure for this model. Browse the model catalog for other checkpoints and their documented download options.
K2 Horizon 0.9B license
The model card sets license: apache-2.0, but its front matter also contains license_name: internal-only and points to a LICENSE file absent from the checked inventory. Those declarations conflict. Confirm the intended license with IFM before commercial use or redistribution; we do not resolve the conflict by copying a sibling model's terms.
Sources and file inventories checked September 15, 2026.
