189 articles covering AI tools, models, and benchmarks.
AI NewsSTADLER Anlagenbau, a 235-year-old German waste recycling equipment maker, has rolled out ChatGPT Enterprise to every...
AI NewsZhipu AI's GLM-5.1 scores 94.6% of Claude Opus 4.6's coding performance in testing. Built on GLM-5's open-source...
ComparisonsBoth cost $10K. Both run Qwen3.5 397B locally. But a dual DGX Spark setup and a Mac Studio M3 Ultra 256GB deliver...
AI NewsOpenAI just launched a Safety Bug Bounty program on Bugcrowd that rewards researchers for finding agentic...
AI NewsGoogle just dropped Gemini 3.1 Flash Live — a real-time audio AI model with 2x longer conversation tracking, 90+...
AI NewsOpenAI releases gpt-oss-safeguard, a free open-source toolkit with prompt-based teen safety policies covering five risk...
BenchmarksA new book by Moritz Hardt argues that benchmark rankings — not scores — are what actually matter. We tested his thesis...
TutorialsSet up the Claude desktop app from scratch — MCP extensions, Cowork agent, Computer Use, and power-user tips that'll...
AI NewsOpenAI Japan just launched its Teen Safety Blueprint — a framework combining age estimation, parental controls, and...
ComparisonsKrasis LLM Runtime claims dramatically faster inference than llama.cpp for large MoE models on a single NVIDIA GPU. We...
BenchmarksATLAS, a source-available AI system built by a Virginia Tech student, scores 74.6% on LiveCodeBench using a single $500...
AI NewsGoogle Lyria 3 is now available to developers through the Gemini API at $0.04 per 30-second clip. Here's what you get,...