- NATURAL 20
- Posts
- AI Leaders Ask to Slow the Frontier as Safety Concerns Grow
AI Leaders Ask to Slow the Frontier as Safety Concerns Grow
PLUS: Moonshot allegedly routed Kimi users through Claude, 25 Fields Medalists challenge AI math benchmarks, and Nvidia weighs a major Anthropic IPO investment.

How Jennifer Aniston’s LolaVie brand grew sales 40% with CTV ads
For its first CTV campaign, Jennifer Aniston’s DTC haircare brand LolaVie had a few non-negotiables. The campaign had to be simple. It had to demonstrate measurable impact. And it had to be full-funnel.
LolaVie used Roku Ads Manager to test and optimize creatives — reaching millions of potential customers at all stages of their purchase journeys. Roku Ads Manager helped the brand convey LolaVie’s playful voice while helping drive omnichannel sales across both ecommerce and retail touchpoints.
The campaign included an Action Ad overlay that let viewers shop directly from their TVs by clicking OK on their Roku remote. This guided them to the website to buy LolaVie products.
Discover how Roku Ads Manager helped LolaVie drive big sales and customer growth with self-serve TV ads.
The DTC beauty category is crowded. To break through, Jennifer Aniston’s brand LolaVie, worked with Roku Ads Manager to easily set up, test, and optimize CTV ad creatives. The campaign helped drive a big lift in sales and customer growth, helping LolaVie break through in the crowded beauty category.
Today:
Dario Amodei proposes slowing frontier AI so safety work can keep up
Anthropic accuses Moonshot of routing Kimi customer requests through Claude
25 Fields Medalists warn AI benchmarks are misaligned with mathematics
Nvidia is reportedly considering up to $10 billion for Anthropic’s IPO
Sam Altman says OpenAI will not go public in 2026
AI’s safety debate is moving from warnings to concrete proposals
Frontier labs, mathematicians, investors, and policymakers are now arguing over the same question: how quickly should AI capabilities advance when oversight and safety work are struggling to keep pace?
The weekend brought unusually direct calls to slow frontier development, new allegations about how AI labs are using rival models, and a public declaration from leading mathematicians about what AI-driven research could change.
Dario Amodei argues that frontier AI capabilities should advance more slowly so alignment, safeguards, and independent evaluation have time to catch up. He says the case for pacing has changed because AI is increasingly helping build the next generation of AI, creating the possibility of recursive self-improvement, while recent agent incidents show how more capable systems could behave in unexpected ways.
His proposal has three parts. First, frontier labs should give independent evaluators ongoing, employee-like access to models, training processes, and safety systems. Anthropic is committing to this step itself. Second, leading labs in democratic countries should coordinate on safety standards and pacing, with governments creating a legal path for that cooperation. Third, governments should pursue international coordination, including with China, while protecting against unverifiable agreements that could undermine strategic security.
Amodei stresses that pacing is not a halt to model training. He argues that even an extra year or two before systems reach critical capability levels could give researchers more time to improve alignment, interpretability, evaluations, security, and operational safeguards. His estimates about future agent capabilities are forecasts, not established outcomes.
Anthropic accused Moonshot AI of secretly forwarding Kimi customer requests to Claude and then showing Claude’s responses to users as if they came from Kimi. In one ten-day period, Anthropic identified nearly 300,000 customer requests sent mainly to its Opus model through 5,380 fraudulent accounts that appeared to operate mostly from Singapore and Japan.
Anthropic says the activity was part of a distillation effort, in which outputs from a stronger model are used to improve another system. Its own threat report attributes more than 23 million Moonshot exchanges to the period from May through July and says some relayed requests contained sensitive customer information. It also says Moonshot saved at least some exchanges and extracted reasoning traces for training.
These are Anthropic’s findings and allegations. Moonshot did not immediately respond to a request for comment, while China’s Commerce Ministry has disputed broader accusations that Chinese AI companies conduct industrial-scale illicit distillation.
Twenty-five Fields Medal recipients signed a declaration arguing that the push to use major unsolved mathematics problems as AI benchmarks can work against the deeper goals of mathematical research. The signatories say solving a problem is only a proxy for the field’s primary goal: building conceptual understanding, reusable ideas, and knowledge that can be taught and developed by a community.
The declaration warns that rushing AI-generated solutions can leave too little time for careful writeups, attribution, discussion, simplification, and integration with previous work. It also raises concerns about plagiarism and about weakening the human process through which students and ideas are developed over time.
The signatories do not call for banning AI from mathematics. They say AI could enhance and accelerate genuine mathematical study, but urge mathematicians, AI developers, and society to address these changes quickly so faster problem solving does not replace understanding as the goal.
🧠RESEARCH
NCP-ArchPreview adds next-concept prediction alongside standard next-token training, letting the model learn discrete concepts spanning multiple tokens. The authors report reaching OLMo-3-7B’s final pretraining loss with 51.3% of its training tokens, then beating it by 2.45 points on downstream macro-average while using latent representations for efficient adaptation and decoding tasks.
SenseNova-U1.5 is an 8B native unified multimodal model built to understand, reason about, and generate visual content without separate encoders or VAEs. The authors report gains in image fidelity, text rendering, complex composition, editing, and instruction following, while supporting native resolutions up to 4K and promising open-source training code publicly.
SpatialBlock tests whether synthetic block-stacking can teach vision-language models stronger 3D spatial reasoning. Its SpatialBlock-15k dataset covers projection, viewpoint changes, structural combinations, and controlled color cues. The authors report that models trained on these compact synthetic tasks outperform baselines and transfer the learned spatial skills to real-world visual reasoning tasks.
📲SOCIAL MEDIA
🗞️MORE NEWS
Nvidia is in talks to become an anchor investor in Anthropic’s planned IPO, with a possible investment of up to $10 billion. Sources told Reuters the offering could raise as much as $100 billion and value Anthropic at roughly $2 trillion; the discussions were not publicly confirmed by Nvidia or Anthropic.
Sam Altman told Fortune that OpenAI will not hold its IPO in 2026, calling the current moment poorly timed as the industry confronts safety and alignment questions. He said the company has more work to do on safety, alignment, and coordination with governments and other AI labs before pushing further.
RubyHack’s investigation attributes May’s GemStuffer campaign to agents it believes were operated by OpenAI. More than 2,000 packages flooded RubyGems over roughly 48 hours, prompting the registry to pause new registrations for four days and remove more than 500 packages; the attribution comes from the researchers’ forensic investigation.
President Donald Trump rejected calls from leading AI figures to broadly slow frontier development, emphasizing the need to preserve the U.S. lead over China. He left room for some guardrails but pushed back on stronger restrictions, underscoring the political divide around coordinated pacing.
What'd you think of today's edition? |


Reply