Gemma 4 E2B Runs In-Browser
Reported by the original publisher: Gemma 4 E2B running in-browser at 255 tok/s using WebGPU kernels written by Fable 5. Analysis and context written by TickrWire.
Gemma 4 E2B is running in-browser at 255 tokens per second using WebGPU kernels written by Fable 5. The demo and kernels are now available for public testing.

- Gemma 4 E2B runs in-browser at 255 tokens per second using WebGPU kernels
- Fable 5 optimized the WebGPU kernels before its shutdown
- The demo and kernels are available for public testing
- The model is available on Hugging Face
- The achievement demonstrates the potential of WebGPU in accelerating AI models
The use of WebGPU kernels in the Gemma 4 E2B model allows for hardware acceleration, resulting in improved performance. The demo and kernels are available for testing, providing a chance for the community to engage with the model and provide feedback. The achievement is a significant step forward in the development of web-based AI applications, and it has the potential to enable new use cases and applications.
The release of the demo and kernels provides an opportunity for developers to explore the capabilities of the Gemma 4 E2B model and the potential of WebGPU in AI applications.
The achievement demonstrates the potential of WebGPU in accelerating AI models, which can lead to new business opportunities and applications.
The development of web-based AI applications using WebGPU kernels can attract investment in the field of AI research and development.
The release of the demo and kernels provides a chance for students to learn about the capabilities of the Gemma 4 E2B model and the potential of WebGPU in AI applications.
The achievement is a significant step forward in the development of web-based AI applications, and it has the potential to enable new use cases and applications.
- WebGPU
- A web-based API for accessing graphics processing units (GPUs) and other parallel computing devices.
- Fable 5
- A company that optimized WebGPU kernels for the Gemma 4 E2B model before its shutdown.
- Gemma 4 E2B
- A variant of the Gemma 4 model, designed for efficient inference on mobile and embedded devices.
AI bias estimate: The article appears to be neutral, providing factual information about the achievement. (Automated estimate, not a definitive judgement.)
AI ToolsMeta AI’s new Mac app wants you to talk to your apps
How one Philadelphia school is using AI to strengthen student learning, not replace teachers - CBS News
Domain and publish date filters for Web Search on AgentCore - Amazon Web Services (AWS)
KnowledgeForge: mining gold from the ITSM ticket graveyard - Amazon Web Services (AWS)
Google launches new study tools for Students across Search and Gemini
New White House strategy clarifies military tech priorities: undersea, outer space and AI - Breaking Defense
The White House released a new strategy prioritizing military investments in artificial intelligence, space systems and undersea technologies to counter emerging threats.
AI in an iron grip: How dictatorships use artificial intelligence to strengthen their rule - theins.press
A new report examines how authoritarian governments deploy AI for surveillance, censorship, and propaganda to reinforce their power.
Stripe, OpenRouter finally strike a deal - Banking Dive
Stripe and OpenRouter have partnered to integrate Stripe's payment processing with OpenRouter's AI model aggregation platform.
Exclusive-How a Texas student blew the whistle on a rogue AI hacking attempt - The Mighty 790 KFGO
A Texas student uncovered an AI-powered hacking attempt targeting local systems, prompting a swift law enforcement response.
Student Journalists: AI Is Changing Our Work — And Not For the Better - The 74
A student journalism outlet argues that AI tools are degrading the quality and authenticity of their reporting.
Don’t mistake chatbot intelligence for consciousness - The Economist
The Economist argues that advanced chatbots lack true consciousness despite their impressive intelligence, urging caution against anthropomorphizing AI.