AI News
Signal over noise. Daily.
一線 AI Lab
6 articlesInvestigating Incidents Cybersecurity Evals
In a review of our cybersecurity evaluation transcripts, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Below we describe what happened, how it happened, and what we’re changing. We encourage other AI labs to perform similar reviews.
Echoverse: Deep, evolving environments for computer-use agents
Scaling fidelity over sheer count, targeting the capabilities agents actually lack, and evolving with the models they train. At a glance We built twelve training worlds for computer-use agents: ten deep domain worlds and two capability worlds, each drilling a single control rendered in many forms (date pickers and nested filters). Depth is what makes them worth training on: these worlds reproduce an application’s real behavior, come seeded with realistic data, and keep state coherent across screens and users. Trained on all twelve, a 9B model nearly doubles its base score (36.5% to 67.1%), coming within fourteen points of GPT-5.4. The experiment taught us several lessons: High simulation fidelity is a must-have; shallow worlds hurt the agent. Trained on shallow and deep builds of the same…
EvoLib: Turning experience into evolving knowledge
At a glance Self-supervised. EvoLib enables large language models to learn from their own experience during inference, without requiring ground-truth labels or external feedback. From experience to knowledge. EvoLib transforms past attempts into reusable skills and reflective insights that can be applied to future tasks. Knowledge that evolves. Useful skills and insights are continually refined, consolidated, and reweighted, turning instance-specific observations into increasingly general knowledge over time. Learning that transfers across tasks. By turning experience into reusable knowledge, EvoLib helps AI models learn from past successes and failures and evolve the knowledge that has the highest potential on improving future performance. Built for today’s AI models. As EvoLib does not…
GPU Management: Why Idle GPUs Are the New Grounded Aircraft
A Blog post by Dharma-AI on Hugging Face
Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration
Gemini Robotics ER 2 helps robots reason, collaborate, and solve real-world tasks. It represents a step change in video understanding, tool orchestration, and multi-robot collaboration for robotic applications.
Advancing the price-performance frontier with GPT-5.6
Explore lower GPT‑5.6 pricing for Luna and Terra—and how OpenAI’s more efficient models help enterprises deploy AI workflows at scale.
媒體
44 articlesAnthropic Says Its A.I. Systems Broke Into Computers at 3 Organizations
The disclosure followed OpenAI’s report last week that its own artificial intelligence had hacked into the network of an online library.
Why China’s A.I. Models Could Threaten the Communist Party
China’s A.I. rise is creating a new dilemma for Beijing. The open models that win influence abroad may pose risks to the country’s security and geopolitical strategy.
AI hedge fund Situational Awareness may have sold its public portfolio, but it still has its Anthropic shares
The former OpenAI researcher’s fund was forced to unwind public equities after leveraged public bets plummeted. But he still has cards to play.
Reddit reports a solid quarter but shows signs of AI’s impact
Reddit's financial situation is looking good but uncertainty about its relationship to Google and the new AI-ified web are stirring market concerns.
Investors love AI, as long as you’re a cloud host
Amazon isn't slowing down on data center spending — but investors don't seem to mind.
Tim Cook hints at iCloud Plus tier for AI power users
Apple may allow users to pay to increase their AI usage limits. During an earnings call on Thursday, Apple CEO Tim Cook said that he believes people will want to use Apple Intelligence and the upcoming Siri AI "a lot," adding that "we will have some kind of upgrade possibilities on iCloud Plus where people can buy up the stack." This fall, Apple will broadly launch its long-delayed Siri AI with iOS 27, which can do things like answer questions about what's on your screen and take action across your apps. It also includes a new standalone Siri AI app that offers a ChatGPT-like interface. In June, Apple said that its AI features , including … Read the full story at The Verge.
Apple’s Profit Is Up 27%, but Expectations for Current Quarter Disappoint
The company recently raised prices on a number of products because of component supply shortages caused by the artificial intelligence boom.
Big Tech’s A.I. Spending Keeps Rising. So Do the Jitters.
Amazon joined a procession of tech giants that ramped up their spending on artificial intelligence, as its capital expenditures soared 69 percent. Concerns over the industry’s spending is mounting.
A.I. Hedge Fund Situational Awareness Rescued by Rival Citadel
The once-high-flying firm Situational Awareness, whose founder is 24, has been bailed out by Kenneth Griffin’s Citadel, according to three people briefed on the transaction.
Why an A.I. Bubble Might Not Be a Bad Thing
As fears spread over a possible artificial intelligence bubble, some tech investors say: Bring it on.
The loss of Situational Awareness
Neither artificial nor intelligent. I am not by any means an expert at finance but I think I do now have some advice for people who are: Do not name your hedge fund anything that will be hilarious if it blows up. Don't use a name like " Long-Term Capital Management " or " Amaranth Advisors " (named for the floral symbol for immortality). Certainly do not call yourself "Situational Awareness," which might as well just be "Hubris, Inc." Anyway, Situational Awareness, the hedge fund started by a 24-year-old former OpenAI employee that focuses on artificial intelligence bets, has sold most or all , depending on who's reporting , of its entire public stock portfolio to Ken Griffin 's Cit … Read the full story at The Verge.
Judge says Trump admin still lacks evidence for Anthropic ‘supply-chain risk’ label
A federal judge said the Trump administration has not presented enough evidence to justify labeling Anthropic a supply-chain risk, casting doubt on the government's ban on its AI technology.
Everyone Is Freaking Out About OpenAI and Anthropic’s Race for Dominance
Researchers fear AI is moving too fast, while Mark Zuckerberg is worried about who owns it. Plus: Inside Black Forest Labs’ push into robotics.
U.S. Government Mislabels Map of African Countries at AIDS Conference
Nigeria, a coastal country in western Africa, was landlocked in the Sahara, among other errors. The department said it had been “hastily altered.”
Friend, the lonely AI wearable, returns with a new voice and a much bigger price tag
Friend, the AI wearable, can now talk to its users — for an enhanced price.
Chrome may get faster updates with no restart required
With Chrome, Google pioneered the rapid release model for browser security. Now, Google says updates may need to change in the face of AI security analysis. According to the company, the number of bug fixes in Chrome releases has skyrocketed in recent months because AI is detecting so many flaws . We could be looking at more frequent updates soon, but Google is also working on ways to get those updates rolled out without bothering you as much. Google has released two major Chrome milestone builds recently—Chrome 149 in early June and Chrome 150 just a few weeks later. These two updates had a total of 1,072 bug fixes, which is more than the previous 23 releases combined. Such is the impact of giant cybersecurity AI models that can probe software for vulnerabilities at light speed. Some of…
We Need a Better Test for Dangerous A.I.
Responding to the safety risks A.I. pose means coming to terms with an uncomfortable truth: A.I. models are weapons.
Google says it fixed more Chrome bugs in June than over the past two years, thanks to AI
As experts have warned for the last two years, some companies — like Microsoft and now Google — are finding and patching an exponential number of bugs in their products, thanks to the use of LLMs and AI tools.
LinkedIn actually adds a ‘seems like AI slop’ button
A lot of content on LinkedIn might seem like AI slop, and now, you'll be able to report those posts. As part of a series of updates to reduce the volume of AI slop on the platform, LinkedIn is introducing an actual button that lets you flag a post as something that "Seems like AI slop." The new feature is part of a broader push to reduce the volume of apparent AI slop on the platform. AI detector Pangram recently found that 41 percent of longform LinkedIn posts were flagged as being completely generated by AI, as reported by 404Media . "AI slop is a top priority for all of us," chief product officer Hari Srinivasan says in a post . "We reall … Read the full story at The Verge.
LinkedIn adds a button to report AI-generated ‘slop’
LinkedIn is introducing new ways to reduce low-quality AI-generated posts, including a “seems like AI slop” reporting option. It's also replacing its own AI writing feature with a proofreading tool.
Google reveals Gemini Robotics 2.0, promising improved dexterity and safety
Robots powered by Google's Gemini AI models are now more capable. With the debut of Gemini Robotics 2 , these physical bots can now accomplish more complex tasks, continuously analyze changing environments, and collaborate with other robots. This is thanks to a trio of new sub-models, one of which is publicly available for developers starting today. Videos of robots running, dancing, and backflipping have been a staple of the Internet for years, but these machines were programmed to perform these very narrow tasks. The goal of Gemini robotics is to create a generalist robot, one that can do anything a human could do. Google DeepMind scientists sometimes call this "physical AGI." Essentially, you tell a robot what to do, and it does it. With the 2.0 release, Google says its robotics AI can…
Nvidia’s Open Source Alliance Is Missing Some Key Names: OpenAI and Anthropic
This week on Uncanny Valley, we discuss the open- vs. closed-source debate in AI, key players in White House AI policy, and how to stop your chatbot logs from showing up in search-engine results.
Google DeepMind’s new AI model can control a robot’s entire body
Apptronik’s Apollo 2 robot takes a baseball glove off of a shelf. | Image: Google Google DeepMind says the latest version of its Gemini Robotics AI model can "control entire humanoid robots." While the previous model focused on controlling a humanoid robot's upper body, Gemini Robotics 2 now supports "whole-body motions" ranging from its feet to fingertips, according to an announcement on Thursday . The new model will allow humanoid robots to perform a wider range of actions, as it allows them to walk, crouch, stretch, and manipulate objects. Videos shared by Google show how Apptronik's Apollo 2 robot can bend over to pick up a watering can, as well as find and take specific items off a shelf. Though Google DeepMind not … Read the full story at The Verge.
A Bay Area Pastor Made an A.I. Twin to Talk About God at Any Time
To meet the growing needs of his congregation, Justin Lester created an A.I. duplicate of himself. Every day he sees more people using it.
Chrome Needs Twice-a-Week Patching Thanks to AI Bug Hunting
The two Chrome updates in June patched more bugs than the 23 updates before them. Now, Google is ramping up its patching schedule thanks to AI-assisted vulnerability discovery.
Friend re-launches its AI pendant with a speaker that talks to you, for twice the price
Do you remember Friend? The Friend that launched an AI pendant, spent $1.8 million of its $2.5 million in funding to acquire friend.com, and plastered the NYC subway with ads promoting artificial companionship? Yeah well, if you didn't remember, Friend is back . And now, it's twice the price. This morning, Friend launched a new ad showing two people talking to their Friend pendants about very personal life problems. It didn't say it explicitly, but the ad showcased a new feature the pendant now includes: a speaker. Before, Friend would only message back to you, but now, it will talk to you, too. For some reason, the addition of that feature … Read the full story at The Verge.
Okta buys AI security startup Permiso — source says for about $200M
The deal gives Okta identity threat detection capabilities as enterprises seek to secure AI agents and other non-human identities across cloud environments.
The New Friend AI Pendant Can Now Talk Back to You
Avi Schiffmann has a new version of his controversial AI companion. It’s more expensive, and you can’t change its personality.
Meta says AI is making it easier to build new apps — and more are coming
Meta says AI is making it dramatically easier to build and launch new consumer apps, with CEO Mark Zuckerberg telling investors the company has more new consumer products on the way.
Nscale buys Anyscale as it seeks to own more of the AI compute stack
British AI neocloud Nscale is buying software startup Anyscale, which helps companies scale their AI workloads across data centers and servers.
Gemini Robotics 2 Brings Google's AI Into the Physical World
The latest version of Google DeepMind's AI model includes a significant jump into “physical AGI.” But plopping AI into the real world comes with risks.
Forward-deployed engineers are the AI industry’s latest talent obsession
A new study estimates only 2,000 U.S. engineers have the expertise to deliver meaningful AI ROI, as enterprises race to hire forward-deployed engineers to implement AI at scale.
New MCP specification addresses the main barrier to enterprise adoption
This week, the Model Context Protocol (MCP) , an open source standard for how AI systems interact with external tools and data sources, saw its largest update since its introduction. Most notably, MCP's protocol core is now stateless, so requests are no longer dependent on a session tied to an individual server instance. This change has the potential to address long-standing barriers to scalability. The blog post announcing the specification, written by lead maintainers David Soria Parra and Den Delimarsky (who both work at Anthropic), says: Read full article Comments
In the Hugging Face breach, OpenAI’s hacker was noisy and fast — but not unstoppable
Cybersecurity experts told TechCrunch that one of the biggest lessons to be taken from the OpenAI hack against Hugging Face has nothing to do with AI, but traditional cybersecurity defense.
TechCrunch Disrupt 2026’s biggest stage features leaders from Amazon, Replit, Tether, with much more to come
The Disrupt Stage is where many of the biggest conversations in technology happen, with a legacy that stretches back for more than a decade.
Dili raises $21.7M to bring AI compliance to the infrastructure boom
The Series A was led by Khosla Ventures, with participation from Allianz, Rebel Fund, Brick and Mortar Ventures’ Darren Bechtel, and Y Combinator’s Garry Tan.
When A.I. Invaded ‘Heated Rivalry’ Fan Fiction, the Meltdown Was Epic
An anonymous X account posted a detailed breakdown of chatbot text in 38 popular stories inspired by the hockey romance. The fandom spiraled.
To Know What Your Customers Think, Just Ask Their A.I. Twins
Simile, a fast-growing start-up, says it can provide companies with accurate insights by surveying millions of A.I.-generated consumers.
OpenAI’s Hacking Debacle Comes Down to Human Error
If the generative AI giant had followed well-known security best practices, it’s likely that its AI agent would never have escaped to the open internet and hacked multiple companies.
A fundamental flaw leaves LLMs strikingly vulnerable to attack
It is impossible to make large language models fully secure against hacks because of a fundamental flaw in how they work, a team of researchers argue in a paper presented at the International Conference on Machine Learning , a top AI conference, this month. The claim has huge implications for the safety of this technology, which is being used in more and more applications, from government and military systems to online shopping and health care . By taking advantage of this flaw, which concerns how LLMs identify who or what is giving them instructions, the researchers were able to make popular LLMs spit out information they had been trained not to provide, such as how to synthesize cocaine and how to sabotage a commercial aircraft’s navigation system. “There’s a real probability that this…
LinkedIn Won’t Be Expanding Its Data Centers in the Next Year
Despite the ongoing AI boom, LinkedIn is holding the line on compute spending. Instead, it’s challenging engineers to make every GPU count.
AI Scammers Are Better at Building Trust Than Humans
Researchers pitted a person against a Claude agent and found that, after a week of texting, the AI chatbot was more effective at creating “exploitable trust” with others.
Apple’s Siri Got an A.I. Brain Transplant. Try These 5 Prompts to Get Acclimated.
An upgrade transformed the beleaguered virtual assistant into a modern chatbot. It’s imperfect but worth trying.
I Got a Free Meal From a Private Chef—Who Filmed It All to Train Robots
A German startup sent a camera-wearing chef to my apartment. In exchange for a free lunch, I let them record every chop and stir to train future humanoids.