- NATURAL 20
- Posts
- AI Agents Move Into Creative Work and Messages
AI Agents Move Into Creative Work and Messages
PLUS: Alibaba launches a $10.2 billion AI-funded share sale, while Hugging Face reportedly explores a sale above $13 billion.

Build. Break. Fix. Learn.
KodeKloud gives you 1,280+ hands-on labs where you provision Kubernetes clusters, write Terraform configs, build CI/CD pipelines, configure Linux systems, containerize apps with Docker, automate with Ansible, and manage Git workflows.
78+ playgrounds let you experiment freely in sandbox AWS environments, Kubernetes clusters, and CI/CD systems without risk.
190+ courses across DevOps, Cloud, and AI pair theory with hands-on labs at every step.
KodeKloud Engineer and 100 Day Challenges provide real-world job scenarios with automated grading that confirms your solutions work.
Stuck? The 55,000+ member Discord community connects you with peers and instructors ready to help.
Every lab runs in a live environment. You deploy, you troubleshoot, you learn. No videos without context. No simulations. The kind of practice that actually builds confidence because you've done real work, not watched someone else do it.
Today:
Spline: V2 lets coding agents edit live 3D scenes
OpenAI: GPT-Image-2 adds transparent backgrounds for reusable assets
OpenAI: Apple Messages plugin lets ChatGPT search and send Mac texts
Alibaba: $10.2 billion share placement funds full-stack AI
Hugging Face: Sale process reportedly targets a $13 billion-plus value
AI Agents Move Into Creative Work and Messages
AI is shifting from generating standalone answers to operating inside the tools where people design, create assets, and communicate.
This edition’s product releases push AI into more consequential workflows. Spline lets agents change a live 3D scene, OpenAI’s image model can produce reusable transparent assets, and ChatGPT can work with private message histories on a Mac.
That usefulness comes with a new control problem. When an agent can edit a project or send a message, approvals, permissions, and clear limits matter as much as model quality.
Spline rebuilt its browser-based 3D editor on WebGPU and added an agent that can inspect a scene, use the editor’s tools, and take screenshots to check its work. It can change objects, materials, lights, cameras, particles, variables, events, and states rather than merely generating a flat preview.
The agent can also write HTML and JavaScript for interfaces, game logic, and interactions. A new MCP server—the standard that lets AI tools connect to software—allows external agents including Claude Code, Cursor, ChatGPT, and Codex to drive the same live, editable scene.
V2 adds separate Edit, Code, and Preview modes, plus physically based materials, HDR lighting, reflections, and updated modeling tools. Spline says WebGPU reduces rendering overhead and keeps heavier scenes interactive, but it does not publish comparative performance tests in the announcement.
The release is available now through Spline’s web and desktop apps. The announcement does not specify new pricing or agent usage limits, so teams should check their plan before assuming the AI and collaboration features are included.
The main limitation is reliability: a screenshot-checking agent can still make incorrect design or code changes. Creators should review scene behavior, generated scripts, and exports before shipping them.

OpenAI added transparent-background output for GPT-Image-2 in the API as a preview. Developers can request a PNG with background="transparent", producing an alpha channel that lets the same asset sit cleanly over different colors and layouts.
The practical uses are straightforward: isolated product shots, presentation charts, stickers, icons, and print designs no longer need a separate background-removal step. OpenAI’s published examples generate 1024-by-1024, 1024-by-1536, and 1536-by-1024 PNGs and verify that the files contain fully transparent pixels.
Access requires an OpenAI API key and access to a transparency-capable image model. The cookbook does not announce a separate price for transparency; normal GPT-Image-2 generation charges apply according to image size and quality.
The preview has an important prompt limitation. If the prompt asks for a scene, backdrop, or background color, that instruction can override the transparent-background setting. OpenAI recommends isolating the subject and explicitly requesting transparency, while users should still inspect edges, small text, and product accuracy.

OpenAI released an Apple Messages plugin for the ChatGPT desktop app on Apple-silicon Macs. Inside ChatGPT Work and Codex, it can read and search iMessage, SMS, and RCS conversations, summarize threads, draft replies, and send messages through Apple’s Messages app.
The plugin is available on all ChatGPT plans, but it does not work in regular ChatGPT chats and does not turn Messages into a remote interface for ChatGPT. It also requires sensitive macOS permissions, including Full Disk Access, contacts, automation, and accessibility access.
Outgoing messages require approval by default. Users can grant persistent approval for a conversation, but OpenAI warns that doing so removes the final review step before the agent sends a message as the user. A known issue can also block sending when a task configuration disables approval prompts.
OpenAI says the optional integration runs locally and does not create a full index of a user’s messages. That is a company description, not an independent privacy audit; users should grant the minimum permissions needed and avoid persistent approval in conversations that may contain untrusted instructions.
🧠RESEARCH
Submitted August 20, AI4AI-Bench tests whether agents can improve AI training algorithms across ten frozen research repositories. Twenty-nine configurations from six systems averaged 0.166 on its normalized scale; the best reached 0.250. More reasoning increased algorithm changes, but the strongest system closed less than one-fifth of the remaining performance gap.
Submitted August 20, researchers train a 1.5-billion-parameter model to choose no, short, or long reasoning upfront. On MATH500, it kept accuracy near baseline—78.2% versus 79.6%—while cutting output from 4,796 to 2,811 tokens. On GSM8K, it reduced tokens 76%, suggesting models can dynamically spend compute by difficulty.
Submitted August 20, PolicyGuide converts organizational rules into workflow graphs that steer customer-service agents through required steps. Across airline, retail, and telecom tests, it raised average Pass4 from 0.42 to 0.62, with telecom improving from 0.19 to 0.61. The approach also transferred across OpenAI, Anthropic, and Google models in testing.
📲SOCIAL MEDIA
🗞️MORE NEWS
Alibaba launched an HK$80 billion, or roughly $10.2 billion, Hong Kong share placement and said all net proceeds would support chips, infrastructure, models, and AI deployment. The offering was enlarged after strong demand; the company did not break down planned spending by category.
Hugging Face has reportedly worked with a bank to gauge buyer interest in a sale valuing the open-model platform at $13 billion or more. Hugging Face had not responded, so no transaction, bidder, or valuation is confirmed.
Anthropic launched Claude Academy with courses, tutorials, progress tracking, and badges covering practical AI use. An optional Claude Academy Skill can recommend courses based on a user’s interests and completed learning; Anthropic frames the curriculum around effective, intentional, and safe use rather than product promotion alone.
Twin1 emerged from stealth with a $20 million seed round for agents grounded in a professional’s emails, meetings, documents, and workplace tools. Linklaters, Orrick, and Dechert are reported users; the company says people can review or reject drafted messages, but its privacy and performance claims have not been independently tested.
What'd you think of today's edition? |



Reply