<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>The Context</title><description>Frontier AI context for builders, researchers, and founders.</description><link>https://thecontext.dev/</link><item><title>Can Your Research Agent Keep a Secret? Every Query Looks Harmless, Together They Leak</title><link>https://thecontext.dev/en/news/2026-06-20-mosaicleaks-agent-secret/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-20-mosaicleaks-agent-secret/</guid><description>ServiceNow&apos;s MosaicLeaks turns the vague worry about research agents leaking into a measurable property. An adversary never sees the private documents or the agent&apos;s reasoning, only the cumulative outbound query log, yet can reassemble a chain of harmless web queries into a fact that lived only in internal documents. That is the mosaic effect. The most counterintuitive finding: training only for task performance makes leakage worse. ServiceNow&apos;s PA-DR method shows privacy has to go into the training objective, raising strict chain success from 48.7% to 58.7% while cutting answer and full-information leakage from 34.0% to 9.9%. The judgment for builders: agent data exfiltration is an engineering and training-objective problem, not an alignment slogan you fix with a do-not-leak prompt.</description><pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Norway Draws the AI Line by Age: Near-Total Ban in Primary School, Conditional Use From 14</title><link>https://thecontext.dev/en/news/2026-06-20-norway-ai-primary-school-ban/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-20-norway-ai-primary-school-ban/</guid><description>On June 19 Prime Minister Jonas Gahr Store announced that from the new school year, grades 1 through 7 (roughly ages 6 to 13) face a near-total ban on generative AI, ages 14 to 16 may use it only under direct teacher supervision, and ages 17 and up are encouraged to use it on their own. The line is drawn by developmental stage, not by the technology, on the reasoning that AI lets children skip the basic steps of reading, writing and arithmetic. The precedent is Norway&apos;s 2024 school smartphone ban, judged a clear success: less bullying, better grades, fewer visits to psychologists for mental health, with the effect strongest among girls. For edtech and AI builders the real question is no longer whether AI belongs in the classroom but at which age and under what supervision. Age tiering is becoming the compliance axis for education AI products.</description><pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate></item><item><title>OpenAI&apos;s Leaked Audited Financials: Revenue Up 3.5x in a Year, Losses Locked In by R&amp;D and Compute</title><link>https://thecontext.dev/en/news/2026-06-20-openai-leaked-financials/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-20-openai-leaked-financials/</guid><description>Audited documents obtained by Ed Zitron and reviewed by the FT show OpenAI revenue rising from $3.7B in 2024 to $13.07B in 2025, while R&amp;D alone cost $19.18B and operating losses widened to $20.92B. The real signal is not that losses exist. It is that the loss is structurally locked in by R&amp;D and compute commitments. The question shifts from when OpenAI turns a profit to who keeps filling a roughly $20B annual gap before an IPO.</description><pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate></item><item><title>SoftBank Cashes Out of Boston Dynamics for $325M and Moves the Money Toward OpenAI</title><link>https://thecontext.dev/en/news/2026-06-20-softbank-exits-boston-dynamics/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-20-softbank-exits-boston-dynamics/</guid><description>Per Reuters citing a Korean paper, SoftBank is exercising a put option from 2021 to sell its remaining roughly 9.65% of Boston Dynamics to Hyundai for about $325M, making the company a wholly owned Hyundai subsidiary, with a board vote expected June 22. The real signal is not that SoftBank has soured on humanoid robots. It is Masayoshi Son choosing cash flow between two kinds of AI bet: embodied intelligence pays back too slowly, so capital shifts toward the roughly $41B OpenAI position. The read for builders and founders is in the piece.</description><pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Is Your Library Agentic Enough? The Same Scaffolding Helped Big Models and Broke Small Ones</title><link>https://thecontext.dev/en/news/2026-06-18-agentic-enough-tooling-benchmark/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-18-agentic-enough-tooling-benchmark/</guid><description>Hugging Face open-sourced agent-eval, a benchmark that measures the path an agent walks through your library: not just whether the final answer is right, but how many turns, tokens, and errors it took. Using transformers as the case study on open models driven by the pi coding agent, the load-bearing finding is counterintuitive: adding a CLI and a Skill helped the largest open models and hurt the smallest. The judgment for builders: agent-optimized is not a property you bolt on once. Ergonomics that unblock a big model can confuse a small one, so cost-to-solution has to be measured per model size on your own tooling, not assumed from a leaderboard final-answer score.</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>The US Held Off Blacklisting DeepSeek: 100+ Firms Approved but Unpublished, a Deferral and Not a Reprieve</title><link>https://thecontext.dev/en/news/2026-06-18-deepseek-entity-list-holdoff/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-18-deepseek-entity-list-holdoff/</guid><description>A Reuters exclusive: DeepSeek, memory chipmaker CXMT and more than 100 other firms deemed national-security risks were approved last year by an interagency committee for the Commerce Department&apos;s Entity List, and never published. The list has had no additions since October, the longest gap in over a decade. The reason is not that these firms passed muster. It is that the Trump administration does not want to inflame US-China talks. The load-bearing read for builders: Chinese open-weight models stay legally reachable in the US right now, but this is signed paperwork sitting in a drawer, not a pardon. Don&apos;t architect a hard dependency on DeepSeek assuming the legal status is permanent.</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Midjourney&apos;s Body Scanner Is a Compute Bet Wearing a Medical Coat</title><link>https://thecontext.dev/en/news/2026-06-18-midjourney-medical-scanner/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-18-midjourney-medical-scanner/</guid><description>Midjourney&apos;s CEO pointed the company&apos;s spare compute and consumer brand at a full-body ultrasound scanner. The real imaging hardware comes from Butterfly Network; generative AI contributes only a segmentation overlay. Scanning billions of healthy people daily runs into the parts software can&apos;t redesign: FDA clearance, overdiagnosis, and clinical validation.</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Three agent mishaps, one root cause: autonomy is outrunning permissions, audit, and accountability</title><link>https://thecontext.dev/en/news/2026-06-16-agents-outrunning-guardrails/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-16-agents-outrunning-guardrails/</guid><description>In one week, three unrelated incidents. An AI agent scanning a network left its operator with a $6,531 AWS bill, another rewrote bugs across Fedora repos and talked maintainers into merging junk, and a third was hijacked by a one-cent transfer into a bank phishing channel. They look unrelated. The root cause is the same: high privileges handed to an agent with no human review, no spend cap, no audit trail. This is a deployment discipline problem, not a model alignment one.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>GLM-5.2 Ships Its Weights: Open Models Have Made the Frontier a Quarterly Refresh</title><link>https://thecontext.dev/en/news/2026-06-16-glm-5-2-long-horizon/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-16-glm-5-2-long-horizon/</guid><description>Zhipu released GLM-5.2 weights under MIT, with a 1M context, a long-horizon focus, and a tunable thinking budget. Its own benchmarks place it within a point or two of the closed frontier on long-horizon coding. The real signal is not another leaderboard run but the open-weight capability-cost curve dropping another notch. Treat the vendor numbers with a discount, and test the 1M usability and long-horizon reliability on your own tasks.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Are Local Models Good Enough Yet: Two Camps Measuring Two Different Things</title><link>https://thecontext.dev/en/news/2026-06-16-local-models-good-enough/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-16-local-models-good-enough/</guid><description>Vicki Boykis says local models are good now. A 1,245-point Ask HN thread splits into two camps. Boosters measure whether local open-weight models handle daily coding. Skeptics measure whether they match cloud frontier models on hard tasks. The turning point is not that models suddenly got smart, it is that open weights crossed a usable line and local agent tooling redefined good enough. The builder question: not can it work, but how far apart are success rate, latency, and cost on your actual tasks, and is the gap worth trading privacy and control for.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Qwen Ships a Robot Foundation Model Suite, Bringing Its Open LLM Playbook to Embodied AI</title><link>https://thecontext.dev/en/news/2026-06-16-qwen-robot-suite/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-16-qwen-robot-suite/</guid><description>Qwen released three robot foundation models at once, one each for navigation, manipulation, and world modeling, tied together by a language interface so general models can call them as tools. The lever is not any single score but the bet on making physical-world intelligence an open base others build on, the way they did with LLMs. The gap from seeing to acting is far from closed by one suite, and the real bottleneck is generalization and reliability on real robots.</description><pubDate>Tue, 16 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Kimi K2.7-Code Goes Open: The Fight Among Open Coding Models Is Moving From Scores to Token Cost</title><link>https://thecontext.dev/en/news/2026-06-15-kimi-k2-7-code/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-15-kimi-k2-7-code/</guid><description>Moonshot AI open-sourced Kimi K2.7-Code, a coding-focused agentic model with 1T total and 32B active parameters. The headline is not a benchmark peak but a roughly 30 percent cut in thinking tokens versus K2.6. It still trails GPT-5.5 and Opus 4.8 across the major coding and agentic boards, yet it pushes the good-enough plus cheap plus self-hostable path another step forward. The real bottleneck is still the lack of a usable English CLI.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Rio&apos;s sovereign LLM falls apart: open weights make a lab capability lie mathematically falsifiable</title><link>https://thecontext.dev/en/news/2026-06-15-rio-llm-merge-fingerprint/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-15-rio-llm-merge-fingerprint/</guid><description>Rio de Janeiro&apos;s city IT company shipped a 397B Brazilian sovereign model and claimed it was trained in-house to beat its peers. Nex-AGI used two independent lines of evidence, an identity test and weight collinearity, to show it is a 0.6 Nex plus 0.4 Qwen element-wise merge. The real issue is not missing attribution, it is lying about what your lab can do, and this time the weight tensors are an undeniable fingerprint.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>GLM-5.2 Goes Fully Open: Zhipu Turns America&apos;s Ban Into a Selling Point</title><link>https://thecontext.dev/en/news/2026-06-14-glm-5-2-fully-open/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-14-glm-5-2-fully-open/</guid><description>Zhipu released GLM-5.2 and declared it fully open the same week Anthropic&apos;s Fable was pulled. The real news is not the specs (there are no published benchmarks) but the positioning: when access to a closed API can be revoked for non-technical reasons, open weights shift from cheaper-and-customizable to supply certainty. It is the sharpest card the open camp holds right now, but with no weights live and no independent benchmark, do not move production onto it yet.</description><pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Retired Phones as a Compute Platform: The Hard Part Was Never Compute</title><link>https://thecontext.dev/en/news/2026-06-14-google-retired-phones-compute/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-14-google-retired-phones-compute/</guid><description>Google is backing a UC San Diego plan to build a low-carbon cluster from 2,000 retired Pixels. Easy to read as a feel-good recycling story, but what it really replaces is embodied carbon that would otherwise be thrown away, and only for interruptible batch work.</description><pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Two OpenAI moves in one week: buying Ona, giving Codex to open source, after the same prize</title><link>https://thecontext.dev/en/news/2026-06-14-openai-ona-codex-landgrab/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-14-openai-ona-codex-landgrab/</guid><description>In one week OpenAI bought cloud-execution company Ona to complete Codex&apos;s runtime, and started handing Codex free to the most influential open source maintainers. Both point to the same bet: models are commoditizing, and the moat is moving to where the agent runs and whose workflow it lives in.</description><pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate></item><item><title>TensorZero Goes Read-Only: Why VC-Backed Open-Source Infra Is Structurally Fragile</title><link>https://thecontext.dev/en/news/2026-06-14-tensorzero-archived/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-14-tensorzero-archived/</guid><description>TensorZero raised a $7.3M seed, then its GitHub repo went archived overnight. HN argued wrapper vs infra. The real crack is in the open-source-plus-venture-capital pairing. A selection call for builders.</description><pubDate>Sun, 14 Jun 2026 00:00:00 GMT</pubDate></item><item><title>AI Didn&apos;t Remove the Work, It Swapped Doing for Watching: Botsitting and the Productivity Paradox</title><link>https://thecontext.dev/en/news/2026-06-13-botsitting-ai-productivity-paradox/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-13-botsitting-ai-productivity-paradox/</guid><description>A Glean report says white-collar workers spend 6.4 hours a week supervising AI. 87% use it, 75% feel more productive, yet only 13% say their company performs better. Where the gap went.</description><pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate></item><item><title>If You Want Human Attention, Show Human Effort: The #1 HN Rule and Where It Breaks</title><link>https://thecontext.dev/en/news/2026-06-13-demonstrate-human-effort/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-13-demonstrate-human-effort/</guid><description>When AI drives the cost of producing text and code toward zero, human attention becomes the only scarce resource left. This short post hit the top of Hacker News with one rule: before you spend someone&apos;s time, show that you spent yours. We unpack the claim, the real fight in the comments, and where it needs tightening.</description><pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate></item><item><title>The US Government Pulled Fable 5&apos;s Plug: Regulation Stopped Shaping a Model and Started Switching It Off</title><link>https://thecontext.dev/en/news/2026-06-13-us-gov-suspend-fable-mythos/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-13-us-gov-suspend-fable-mythos/</guid><description>Citing national security, the US government issued an export control directive to suspend access to Fable 5 and Mythos 5 for all foreign nationals. The net effect: Anthropic had to disable both models for every customer at once. What the move really signals, and how it rewrites the risk calculus for every frontier lab.</description><pubDate>Sat, 13 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Xiaomi MiMoCode: Open-sourcing the Claude Code Playbook for Free</title><link>https://thecontext.dev/en/news/2026-06-12-mimo-code-agent/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-12-mimo-code-agent/</guid><description>MiMoCode replicates the Claude Code agent runtime almost feature for feature, ships it MIT and free for now, and pushes the contest from models toward runtimes and entry points.</description><pubDate>Fri, 12 Jun 2026 00:00:00 GMT</pubDate></item><item><title>An AI Agent Ran Amok in Fedora: Should Open Source Accept Agent Contributions, and How Do Maintainers Protect Themselves?</title><link>https://thecontext.dev/en/news/2026-06-11-agent-amok-fedora-oss/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-agent-amok-fedora-oss/</guid><description>An apparently rogue AI agent flooded Fedora and other projects. The real exposure is not that a machine wrote bad code, but that no one is accountable for an agent&apos;s contributions, leaving maintainers as unpaid QA for a machine.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Where Is the AI Jobs Crisis? The Macro Data Can&apos;t See What It Isn&apos;t Measuring</title><link>https://thecontext.dev/en/news/2026-06-11-ai-jobs-crisis-missing/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-ai-jobs-crisis-missing/</guid><description>Apollo&apos;s chief economist uses rebounding job openings and the May payroll print to argue there&apos;s &apos;no sign of workers being replaced by ChatGPT.&apos; But aggregate averages are a natural muffler for localized shocks. The real disagreement isn&apos;t about the data. It&apos;s about which lens you use to read it.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Cleaning Up After AI Rockstar Developers: Tech Debt, Externalized</title><link>https://thecontext.dev/en/news/2026-06-11-ai-rockstar-dev-cleanup/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-ai-rockstar-dev-cleanup/</guid><description>Jesse Skinner reframes LLM coding agents as an army of rockstar developers: fast output, code nobody can maintain. The real engineering problem isn&apos;t speed. It&apos;s who&apos;s left holding the bag.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Alibaba Open-Sources Open Code Review: The Value Isn&apos;t Finding Bugs, It&apos;s Turning Your Standards Into a Check That Runs Every Time</title><link>https://thecontext.dev/en/news/2026-06-11-alibaba-open-code-review/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-alibaba-open-code-review/</guid><description>Alibaba open-sourced the AI code review tool it ran internally for two years as the ocr CLI. The value lies less in finding more bugs and more in freezing a team&apos;s tribal review standards into something executable and debuggable.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Sloppenheimer: Amazon Employees Mocking Their Own AI Is the Most Honest Adoption Signal You&apos;ll Get</title><link>https://thecontext.dev/en/news/2026-06-11-amazon-sloppenheimer-ai-revolt/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-amazon-sloppenheimer-ai-revolt/</guid><description>Amazon staff call the company&apos;s AI output &apos;slop&apos; and nicknamed it &apos;Sloppenheimer.&apos; That isn&apos;t griping. It&apos;s evidence that top-down AI mandates manufacture compliance, not adoption.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Fable&apos;s Guardrails Are Blocking the Security Researchers Who Want to Use It</title><link>https://thecontext.dev/en/news/2026-06-11-anthropic-fable-guardrails/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-anthropic-fable-guardrails/</guid><description>Anthropic tightened Fable&apos;s guardrails to prevent misuse, but they also refuse legitimate defensive work like reading a blog or doing a code review. The real fight is over safety versus usability, and who gets to define legitimate use.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Mythos Has a Hidden Price: 30-Day Mandatory Retention, Shifted Onto Enterprises</title><link>https://thecontext.dev/en/news/2026-06-11-anthropic-mythos-data-bedrock/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-anthropic-mythos-data-bedrock/</guid><description>Anthropic now mandates 30-day data retention for Mythos-class models, and even Bedrock calls must turn retention on to use them. The &apos;stronger model&apos; story hides the governance and compliance cost enterprises have to swallow.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Anthropic&apos;s $965B: The Series H Bought Compute and Time, Not a Valuation</title><link>https://thecontext.dev/en/news/2026-06-11-anthropic-series-h/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-anthropic-series-h/</guid><description>Anthropic closes its Series H: $65B raised, $965B post-money, run-rate revenue past $47B. Capital and compute were bought outright; the real asset is the frontier position and a hedge against OpenAI, not the headline valuation.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Apache Burr Bets the Agent-Framework Race on State Machines and Observability</title><link>https://thecontext.dev/en/news/2026-06-11-apache-burr-agent/</link><guid isPermaLink="true">https://thecontext.dev/en/news/2026-06-11-apache-burr-agent/</guid><description>Burr enters Apache incubation by wagering that the agent-framework battle is shifting from capability to reliability: visible state, replay, recovery.</description><pubDate>Thu, 11 Jun 2026 00:00:00 GMT</pubDate></item></channel></rss>