Easy AI Daily — January 28, 2026
A structured digest of the day's AI industry news, curated by the Easy AI Daily team.
Key points
- Moonshot released Kimi K2.5: an open-weight 1T-parameter MoE (32B active) multimodal model achieving open-model SOTA on HLE, BrowseComp, MMMU Pro, VideoMMMU, and SWE-bench; native image/video understanding, 128K→256K context, INT4 quantization, runs locally on multi-GPU Macs. Available on HuggingFace, Ollama, Together, Fireworks. Tech blog
- Trinity Large (Arcee × Prime Intellect × Datology): 400B MoE with 13B active parameters, trained on 17T tokens (~2000 B300s for a month), featuring 3:1 local/global gated attention, SWA, NoPE+RoPE, and the Muon optimizer. Day-one vLLM support; free via OpenRouter for now. Announcement
- DeepSeek-OCR 2 open-sourced: introduces Visual Causal Flow and DeepEncoder V2, compressing images to ~256–1120 vision tokens; scores 91.09% on OmniDocBench v1.5 (+3.73). Model
- OpenAI Prism: a free GPT-5.2-powered research workspace combining LaTeX writing, collaboration, and literature management for all personal ChatGPT accounts; no automatic IP claims. Launch
- Transformers v5 released: 6–11x MoE speedups, ~50% faster single-request inference, doubled concurrent throughput. Repo
- Agent ecosystem momentum: Kimi Agent Swarm (up to 100 sub-agents, 1500 tool calls, PARL-trained scheduling), open-source Kimi Code (Apache-2.0) and Agent SDK, Google Jules' Planning Critic (-9.5% task failure), and Karpathy betting on agent-first programming.
- A paper argues LLM hallucinations are mathematically inevitable in the current paradigm, and jailbreaks amplify them: arXiv:2409.05746
- Anthropic shows minimal fine-tuning on frontier-model outputs can restore suppressed dangerous (bio-risk) capabilities: PDF
- DeepPlanning benchmarks long-horizon planning; PrefixRL ~2x faster RL convergence; Google ATLAS multilingual scaling laws; Epoch's FrontierMath remains unsolved by any AI.
- MergeMix uses learnable model merging to optimize data mixtures mid-training: arXiv PDF
- Unsloth: MoE training now ~14x faster than v4, targeting 30x. Post
- FlagOS, tinygrad Megakernels, FlashInfer-Bench datasets released for MLSys 2026 contest: trace data
- Gemini controversy: users report Pro/Ultra actual "hot memory" of only 32k–128k tokens vs. advertised millions, plus quota cuts and billing bugs (one user reportedly overcharged $70k+), pushing users to Grok 4.1 and Claude Sonnet 4.5.
- Clawdbot renamed Moltbot over "Claude" trademark conflict, with community-exposed security issues around unauthorized environment-variable access.
- MiniMax teases M2.2; Qwen hints at new vision models; a16z reports 80% of startups use Chinese open-source models.
- AI text detectors keep mislabeling pre-GPT human papers as AI-generated, causing academic harm.
Research highlights
Infrastructure & products
📌 Source: Easy AI Daily · 🤖 Compiled by: AI assistant