ai newsroom coverage

8 articles

← All topics

OpenAI's Jalapeño chip posts first measured numbers — 1.5–1.9× more AI work per watt and 1.7–3.6× lower end-to-end latency
openaijalapenoinference-chipcustom-siliconbenchmarks+9

OpenAI published first-party benchmark numbers for Jalapeño: 1.5–1.9× more AI work per watt, 1.7–3.6× lower end-to-end latency. Vendor-tested, not independent.

context-mode: 98% context savings and session continuity across 17 AI coding agents
context-modemcpcontext-engineeringcontext-windowai-coding-agents+4

mksglu/context-mode is a 19.4k★ MCP server with 98% tool-output savings, SQLite/FTS5 session-continuity, and 17 supported agent platforms — under an ELv2 license.

OpenAI's GPT-Red: self-play red-teaming at frontier scale; GPT-5.6 Sol 6× more robust
openaigpt-5-6gpt-5-6-solgpt-redautomated-red-teaming+7

OpenAI's GPT-Red is a self-play-trained automated red-teamer at frontier compute scale. GPT-5.6 Sol is 6× more robust to prompt injection; specific Vendy and Codex CLI exploits.

Ratel: an in-process BM25 tool catalog that cuts AI agent context spend 87% on BFCL v3
ratelratel-aicontext-engineeringtool-callingtool-selection+23

Ratel (ratel-ai/ratel, 186★, Apache-2.0 core + MIT SDKs) ships an in-process BM25 tool catalog. On BFCL v3: ~87% fewer tokens, tool selection within ±5 points.

Sonnet 5 launches at $2/$10, nears Opus 4.8
anthropicclaude-sonnet-5claude-opus-4-8anthropic-pricinganthropic-api+20

Anthropic's Claude Sonnet 5 launches 2026-06-30 with intro pricing $2/$10 per MTok (→$3/$15 on Sep 1), a new effort-level API dial, and a near-Opus cost-performance curve.

cognee: open-source AI memory platform for agents
cogneetopoteretesai-memoryagent-memoryknowledge-graph+22

cognee is an Apache-2.0 open-source AI memory platform for agents: a self-hosted knowledge graph engine with a four-method API (remember, recall, forget, improve) and a Claude Code plugin.

OpenAI and Broadcom unveil Jalapeño, OpenAI's first custom LLM inference chip
openaibroadcomjalapenoinference-chipcustom-silicon+13

On 2026-06-24 OpenAI and Broadcom unveiled Jalapeño, OpenAI's first custom LLM inference chip. Lab samples run today; gigawatt-scale deployment with Microsoft is planned for 2026.

codebase-memory-mcp: zero-dep code intelligence
codebase-memory-mcpmcpmodel-context-protocoltree-sitterknowledge-graph+12

Pure-C, single-binary MCP server that indexes a codebase into a Tree-Sitter knowledge graph in milliseconds. 13.3k stars, MIT, 5,604 tests, 11 agents. arXiv 2603.27277 reports 83% quality at 10× fewer tokens.