Skip to content

WorthPosting

  • Home
  • About

Tag: LLM

Cat Links AI News

Training-Inference Mismatch: Why Your LLM Reinforcement Learning Is Optimizing the Wrong Policy

Posted on July 12, 2026July 13, 2026

Reinforcement learning has become the defining ingredient of modern LLM post-training. GRPO, PPO, and their variants drive the reasoning capabilities

Continue readingTraining-Inference Mismatch: Why Your LLM Reinforcement Learning Is Optimizing the Wrong Policy

Cat Links AI News

GLM-5.2 and Tencent Hy3: Two Different Bets on the Open-Weight Frontier

Posted on July 8, 2026July 9, 2026

The open-weight frontier has been moving fast. Over the past few weeks, two major releases have landed on HuggingFace that

Continue readingGLM-5.2 and Tencent Hy3: Two Different Bets on the Open-Weight Frontier

Cat Links AI News

Program-as-Weights: Compiling Natural Language Into Local Neural Programs

Posted on July 5, 2026

There’s a class of programming tasks that resists clean implementation: deciding whether a log line is “important,” repairing malformed JSON

Continue readingProgram-as-Weights: Compiling Natural Language Into Local Neural Programs

Cat Links Software Engineering

vLLM v0.23.0: Model Runner V2, Multi-Tier KV Offloading, and the Growing Rust Frontend

Posted on June 22, 2026June 23, 2026

The vLLM v0.23.0 release landed last week with 408 commits from 200 contributors, and it packs several changes that directly

Continue readingvLLM v0.23.0: Model Runner V2, Multi-Tier KV Offloading, and the Growing Rust Frontend

Cat Links AI News

LoopCoder-v2: Why Two Loops Beat Four in Test-Time Compute Scaling

Posted on June 21, 2026June 22, 2026

The dominant scaling narrative in large language models has been straightforward: more parameters, more data, more compute. But there’s a

Continue readingLoopCoder-v2: Why Two Loops Beat Four in Test-Time Compute Scaling

Cat Links AI News

GLM-5.2: The New #1 Open-Weight LLM and Why IndexShare Matters

Posted on June 17, 2026June 18, 2026

The open-source LLM landscape just got a new heavyweight contender. Z.ai (Zhipu AI) released GLM-5.2, a 753B-parameter mixture-of-experts model that

Continue readingGLM-5.2: The New #1 Open-Weight LLM and Why IndexShare Matters

Cat Links AI News

How MiniMax Sparse Attention Achieves 28x Compute Reduction at 1M Context Length

Posted on June 14, 2026

The attention mechanism is the backbone of every transformer model, but it carries a brutal cost: quadratic complexity with respect

Continue readingHow MiniMax Sparse Attention Achieves 28x Compute Reduction at 1M Context Length

Cat Links Software Engineering

5 Trending GitHub Repos: Apple’s Container Runtime Hits 1.0, LLM Token Compression, and AI Skill Security

Posted on June 13, 2026

The GitHub trending page this week is dominated by AI agent tooling, but tucked between the skills and plugins are

Continue reading5 Trending GitHub Repos: Apple’s Container Runtime Hits 1.0, LLM Token Compression, and AI Skill Security

Cat Links AI News

Microsoft’s MAI Models at Build 2026: Seven New AI Models and What They Mean for Developers

Posted on June 3, 2026June 4, 2026

Microsoft’s Build 2026 conference delivered a move that had been anticipated for months but still landed with weight: the company

Continue readingMicrosoft’s MAI Models at Build 2026: Seven New AI Models and What They Mean for Developers

Cat Links AI News

SkillOpt: Training AI Agent Skills Like Neural Networks

Posted on May 31, 2026June 1, 2026

AI agents have a skill problem. You give a language model a system prompt — or “skill” — and it

Continue readingSkillOpt: Training AI Agent Skills Like Neural Networks

Posts navigation

Older posts
Newer posts
  • Home
  • About
Copyright © 2026 WorthPosting | Signify by WEN Themes
Scroll Up