Researchers Charles Ye, Jasmine Cui, and Dylan Hadfield-Menell have published a paper and accompanying blog post exploring prompt injection as a form of role confusion in large language models. Their work investigates the challenge of models distinguishing between privileged system text and untrusted user input, particularly when using role tags like <system> and <user>. The findings confirm existing difficulties in preventing prompt injection attacks.
→ This research directly addresses a core vulnerability in LLM security, particularly relevant for anyone building RAG systems or local LLMs where robust input handling is critical.
Simon Willison successfully ported Moebius, a 0.2B image inpainting model, to run in a web browser using WebGPU. The original model required PyTorch and NVIDIA CUDA, but the web-based version now allows users to remove regions of an image and have the model intelligently fill the space directly in their browser. A demo is available at simonw.github.io/moebius-web/.
→ This is a prime example of a lightweight, high-performance model being made accessible on-device, aligning perfectly with the goals of projects like LiteRT-LM and MediaPipe.
Hugging Face is experimenting with the proposed Cross-Origin Storage API to enable Transformers.js to download and cache models more efficiently across different origins. This new API aims to improve performance and user experience for web-based AI applications by allowing shared storage for large model files. The current implementation uses a polyfill to demonstrate the potential benefits before the API is widely adopted.
→ This could significantly boost on-device AI performance for web apps using local LLMs like Gemma, making model loading faster and more reliable across domains.
Simon Willison successfully ported Moebius, a 0.2B image inpainting model, to run in a web browser using WebGPU. The original model required PyTorch and NVIDIA CUDA, but the web-based version now allows users to remove regions of an image and have the model intelligently fill the space directly in their browser. A demo is available at simonw.github.io/moebius-web/.
→ This is a prime example of a lightweight, high-performance model being made accessible on-device, aligning perfectly with the goals of projects like LiteRT-LM and MediaPipe.
Hugging Face is experimenting with the proposed Cross-Origin Storage API to enable Transformers.js to download and cache models more efficiently across different origins. This new API aims to improve performance and user experience for web-based AI applications by allowing shared storage for large model files. The current implementation uses a polyfill to demonstrate the potential benefits before the API is widely adopted.
→ This could significantly boost on-device AI performance for web apps using local LLMs like Gemma, making model loading faster and more reliable across domains.
Hugging Face announced a new initiative to use local models for triaging issues in the OpenClaw repository without incurring costs. This leverages on-device processing capabilities to manage open-source project contributions efficiently. The project aims to demonstrate the practical application of local AI for developer tooling.
→ This is a prime example of local LLMs like Gemma or Mistral being deployed for practical, cost-free developer tooling, directly relevant to on-device AI and open-source project management.
Researchers Charles Ye, Jasmine Cui, and Dylan Hadfield-Menell have published a paper and accompanying blog post exploring prompt injection as a form of role confusion in large language models. Their work investigates the challenge of models distinguishing between privileged system text and untrusted user input, particularly when using role tags like <system> and <user>. The findings confirm existing difficulties in preventing prompt injection attacks.
→ This research directly addresses a core vulnerability in LLM security, particularly relevant for anyone building RAG systems or local LLMs where robust input handling is critical.
GPT-5 Pro assisted immunologist Derya Unutmaz in resolving a three-year-old mystery concerning T cell behavior. This breakthrough provides new insights that could advance research in cancer and autoimmune diseases. The AI's role involved processing complex immunological data to identify previously unobserved patterns.
→ While not a local LLM, this highlights the potential for advanced models like GPT-5 to accelerate scientific discovery, hinting at future capabilities for local models in specialized research applications.
Researchers Charles Ye, Jasmine Cui, and Dylan Hadfield-Menell have published a paper and accompanying blog post exploring prompt injection as a form of role confusion in large language models. Their work investigates the challenge of models distinguishing between privileged system text and untrusted user input, particularly when using role tags like <system> and <user>. The findings confirm existing difficulties in preventing prompt injection attacks.
→ This research directly addresses a core vulnerability in LLM security, particularly relevant for anyone building RAG systems or local LLMs where robust input handling is critical.
Simon Willison successfully ported Moebius, a 0.2B image inpainting model, to run in a web browser using WebGPU. The original model required PyTorch and NVIDIA CUDA, but the web-based version now allows users to remove regions of an image and have the model intelligently fill the space directly in their browser. A demo is available at simonw.github.io/moebius-web/.
→ This is a prime example of a lightweight, high-performance model being made accessible on-device, aligning perfectly with the goals of projects like LiteRT-LM and MediaPipe.
Hugging Face is experimenting with the proposed Cross-Origin Storage API to enable Transformers.js to download and cache models more efficiently across different origins. This new API aims to improve performance and user experience for web-based AI applications by allowing shared storage for large model files. The current implementation uses a polyfill to demonstrate the potential benefits before the API is widely adopted.
→ This could significantly boost on-device AI performance for web apps using local LLMs like Gemma, making model loading faster and more reliable across domains.
Hugging Face announced a new initiative to use local models for triaging issues in the OpenClaw repository without incurring costs. This leverages on-device processing capabilities to manage open-source project contributions efficiently. The project aims to demonstrate the practical application of local AI for developer tooling.
→ This is a prime example of local LLMs like Gemma or Mistral being deployed for practical, cost-free developer tooling, directly relevant to on-device AI and open-source project management.
GPT-5 Pro assisted immunologist Derya Unutmaz in resolving a three-year-old mystery concerning T cell behavior. This breakthrough provides new insights that could advance research in cancer and autoimmune diseases. The AI's role involved processing complex immunological data to identify previously unobserved patterns.
→ While not a local LLM, this highlights the potential for advanced models like GPT-5 to accelerate scientific discovery, hinting at future capabilities for local models in specialized research applications.
Tool: OPFS + Pyodide test harness I've been pondering if Datasette Lite - the Python Datasette application run entirely in the browser using Pyodide and WebAssembly - might be able to edit persistent SQLite files stored on the user's computer. That's what OFPS (Origin Private Fil
Learn how Jason Liu uses Codex to preserve context, manage complex projects, and help work continue beyond a single prompt.
We have covered the Age of Async Agents on the podcast: There has been a wave of companies building their own background agents from Shopify to Stripe to Paradigm to Razorpay , and even Cognition’s friends Ramp have built their own coding agent with other friend Modal . And today
AI Engineer World’s Fair regular bird tix will sell out ~today! Join us next week ahead of the Late Bird price hike and get >$40,000 in sponsor credits for attending ! Thanks to the US Government issuing an export control directive on Mythos and Fable , the risks of jailbreaks an
9 free skills and plug-ins that make Claude Code, Codex, and other AI coding agents dramatically more useful, and yes, all of them are free. Skills and plug-ins are reusable instruction files that give your AI agent a repeatable behavior instead of starting from scratch every tim
I messed up. Everyone has been talking about the drama that went down around Claude Fable and Mythos getting shut down, including myself. But I got distracted from the REAL story here, which is just how crazy good these models were and the amazing things people were able to build
In the last few months, I've started to see [job applications] that were clearly cowritten by an LLM, link to an LLM-generated portfolio site, which then links to LLM-generated GitHub projects, with purely LLM-generated commit messages. [...] My other reaction is that I don't kno
Release: datasette 1.0a35 I'll write more about this one soon, but it's a big release. Three highlights from the release notes: New "Create table" interface in the database actions menu, backed by the /<database>/-/create JSON API . It can define columns, primary keys, custom col
A new OpenAI research paper shows how AI agents are transforming work, enabling longer, more complex tasks and expanding productivity across roles.
OpenAI and Broadcom introduce Jalapeño, a custom AI chip built for LLM inference to improve performance, efficiency, and scale across AI systems.
OpenAI introduces Patch the Planet, a Daybreak initiative helping open-source maintainers find, validate, and fix vulnerabilities with AI and expert review.
The brief history of Meta-Harnesses is a little undocumented, but it roughly goes: at first there was Conductor and Zed’s ACP , then there came OpenInspect , Cloudflare’s Flue , and then Vercel’s Eve and HarnessAgent , and Heypi . It should not go unnoticed that today’s podcast g
We’re excited to have Databricks join us at AIEWF , among hundreds of the top companies in the AI Engineer ecosystem. LS subscribers can use their discount to get past the late bird pricing and access over $50k in sponsor offers ! Everyone is still talking about Satya’s Frontier
Congrats due to Baseten, who officially announced their leaked $13B Series F. Today had a smattering of midsize news across OpenAI Daybreak and Gemini Interactions and Sakana Fugu, but probably the trend to watch and hang your hat on is SpaceX’s THIRD GPU rental deal, this time w
Codex can now watch your screen, learn what repetitive tasks you need done, and then do them for you. OpenAI just dropped a new feature called "Record and Replay" and it’s the ultimate cheat code for repetitive tasks. First, you record yourself doing a task once and the AI will w
Did PewDiePie actually build a ChatGPT competitor? Felix dropped $41,000 on an insane local AI rig and built his own open-source AI workspace with zero prior coding knowledge. While the internet is calling it a ChatGPT competitor, it’s actually something much more important: a co
Claude Tag embeds persistent, proactive AI coworkers into Slack channels to access team context, manage long‑horizon tasks, and automate code and incident response. Regulatory flashpoints include the Anthropic Fable access ban, a lawsuit seeking model reinstatement, and governmen
Examination of the Fable/Mythos controversy, including the NSA jailbreak context and Trump’s remarks about Anthropic. Analysis of GLM 5.2 and DeepMind departures showing open-weight models challenging frontier labs and reshaping enterprise AI tactics. Rumors about Fable, Claude S
simonw/browser-compat-db Inspired by Mozilla's new MDN MCP service - source code here - I decided to try converting their comprehensive mdn/browser-compat-data repository full of browser compatibility data into a SQLite database. This new GitHub repo includes a Claude Code for we
OpenAI helps build shared standards for advanced AI, supporting evaluation frameworks, safety practices, and global cooperation through the Appia Foundation.
Discover how Omio uses OpenAI to power conversational travel experiences, accelerate product development, and transform into an AI-native company.
OpenAI introduces new Daybreak tools, including Codex Security and GPT-5.5-Cyber, to help organizations find, validate, and patch vulnerabilities at scale.
tp3_memories_localgemma3:4b)