SDAD Writes

Notes from Sai Dutta Abhishek Dash — AI infrastructure & security engineer — on LLM gateways, code-security agents, self-hosted systems, and the open-source tools he ships.

K2 Horizon: Six Fully Open Models and 21TB of Training Data

IFM released six foundation models from 0.9B to 375B, plus roughly 21.5TB of the actual training datasets, all Apache 2.0. I verified the dataset sizes and benchmarked the three small models myself: the 0.9B does real reasoning at 19 tokens per second in 2.2GB of VRAM. Here is what shipped, what is still missing, and where the catches are.

Qwen3.8-27B Benchmark on NVIDIA DGX Spark GB10

Real Qwen3.8-27B numbers on the DGX Spark GB10: the memory wall caps single-stream decode at 11-12 tok/s, and speculative decoding is the only way out. SGLang with DSpark hits 34 tok/s, veloGB10 reaches about 40 tok/s, and NVFP4 keeps quality within noise of FP8.

Best Google Colab Alternatives Free GPU Tiers Ranked

I ranked nine free GPU notebook platforms against Google Colab in 2026. Kaggle (Tesla P100, 30 hrs/week, 29GB RAM) and AWS SageMaker Studio Lab (T4, 4 GPU hrs/day, persistent storage) take S tier, Paperspace Gradient is A tier with persistent storage, and Colab itself lands B tier for vague limits and 2-3 hour disconnects.