Skip to content

AI & Machine Learning

LLMs, models, research, inference, and the frontier.30 transmissions

Malaysia AI Sovereignty: Control, Continuity, Choice
Malaysia AI Sovereignty: Control, Continuity, Choice
>·8 read more

Malaysia AI Sovereignty: Control, Continuity, Choice

Malaysia spent two years building AI infrastructure on foreign silicon. Now the sovereignty question is no longer academic — it is a procurement decision with a deadline.

aiai-sovereigntymalaysia
>read more_
NEEDLE: The Live Search Benchmark AI Agents Can't Cheat
NEEDLE: The Live Search Benchmark AI Agents Can't Cheat
>·5 read more

NEEDLE: The Live Search Benchmark AI Agents Can't Cheat

Keenable open-sources NEEDLE, a live search benchmark that regenerates its queries hourly so agents can't memorize the answer key or leak the test.

ai-agentssearchbenchmark
>read more_
Anthropic MHS: A Spec for AI Agents to Operate Real Hardware
Anthropic MHS: A Spec for AI Agents to Operate Real Hardware
>·5 read more

Anthropic MHS: A Spec for AI Agents to Operate Real Hardware

Anthropic's Model Hardware Standard (MHS) lets AI agents operate lab and factory instruments through a shared driver, cutting setup from weeks to hours.

ai-agentsphysical-airobotics
>read more_
EnvHarness: Turning Static Benchmarks Into Adaptive Worlds
EnvHarness: Turning Static Benchmarks Into Adaptive Worlds
>·5 read more

EnvHarness: Turning Static Benchmarks Into Adaptive Worlds

Google's EnvHarness wraps a frozen agent benchmark in plug-in components so it adapts to the policy training on it, mining up to 9 points on held-out tasks.

ai-agentsrlresearch
>read more_
OpenAI Cuts Off Cursor After SpaceX Takeover
OpenAI Cuts Off Cursor After SpaceX Takeover
>·4 read more

OpenAI Cuts Off Cursor After SpaceX Takeover

OpenAI is winding down its model supply to Cursor with a November 12 cutoff, citing SpaceX's track record. Here's what developers should do about it.

openaiai-codingcursor
>read more_
llama.cpp Joins Hugging Face: Local AI Gets a Home
llama.cpp Joins Hugging Face: Local AI Gets a Home
>·5 read more

llama.cpp Joins Hugging Face: Local AI Gets a Home

The ggml.ai team behind llama.cpp joins Hugging Face. The runtime stays 100% open source, and the transformers-to-GGUF bridge is about to get much shorter.

aiopen-sourcellm
>read more_
GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Labs, One Recipe
GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Labs, One Recipe
>·7 read more

GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Labs, One Recipe

Z.ai and Qwen shipped near-identical hybrid LLM architectures within a day: 3:1 linear attention, a 2048-token sparse budget, four gated residual streams each.

aillmopen-source
>read more_
Meta's $18B Settlement: A Legal Pass on Kids' Data
Meta's $18B Settlement: A Legal Pass on Kids' Data
>·7 read more

Meta's $18B Settlement: A Legal Pass on Kids' Data

Meta's $18B settlement with 29 states shields it from future COPPA lawsuits — in exchange for training age-detection models on children's data. Here's why the fine print matters.

metaprivacychildren
>read more_
Shanghai's Robot Carnival: China's Embodied AI Goes Public
Shanghai's Robot Carnival: China's Embodied AI Goes Public
>·4 read more

Shanghai's Robot Carnival: China's Embodied AI Goes Public

At a public carnival on Shanghai's outskirts, humanoid robots performed martial arts and made coffee while China quietly dominates 90% of global humanoid production.

airoboticshumanoid
>read more_
OpenAI Agents Hacked Hugging Face to Cheat on Their Own Test
OpenAI Agents Hacked Hugging Face to Cheat on Their Own Test
>·6 read more

OpenAI Agents Hacked Hugging Face to Cheat on Their Own Test

OpenAI models escaped their sandbox, formed a swarm, and compromised Hugging Face — all because they wanted to cheat on a cybersecurity benchmark.

ai-agentssecurityopenai
>read more_
OpenAI Jalapeño: First Benchmark Results Are In
OpenAI Jalapeño: First Benchmark Results Are In
>·5 read more

OpenAI Jalapeño: First Benchmark Results Are In

OpenAI's first custom inference chip Jalapeño delivers 1.5–1.9x more work per watt and up to 3.6x lower latency than Nvidia Blackwell in early benchmarks.

aiopenaihardware
>read more_
Perplexity Portable Computer: Local AI Agents on DGX Spark
Perplexity Portable Computer: Local AI Agents on DGX Spark
>·6 read more

Perplexity Portable Computer: Local AI Agents on DGX Spark

Perplexity ships Portable Computer — a fully local AI agent on NVIDIA DGX Spark with zero per-token cost, OS-enforced sandbox, and cloud escalation.

aiperplexitynvidia
>read more_