Posts tagged “training

1 post on SDAD Writes.

K2 Horizon: Six Fully Open Models and 21TB of Training Data

IFM released six foundation models from 0.9B to 375B, plus roughly 21.5TB of the actual training datasets, all Apache 2.0. I verified the dataset sizes and benchmarked the three small models myself: the 0.9B does real reasoning at 19 tokens per second in 2.2GB of VRAM. Here is what shipped, what is still missing, and where the catches are.