News highlights
- Arm AI Portal connects more than 22 million developers and their agents to optimized AI software across the Arm compute platform spanning cloud, edge and physical AI
- Developers can start faster with pre-optimized models or bring and optimize their own models for best performance on Arm
- As development becomes increasingly agentic, AI Portal makes models, performance data and workflows machine-discoverable, meeting developers and agents where they already build, giving simpler access to Arm’s latest AI technologies
Agentic AI is moving beyond the cloud to edge and physical AI, creating complexity as developers build across models, runtimes and hardware targets. Arm’s compute platform spans this continuum, supported by more than 22 million developers who need optimized AI software.
Building an AI application shouldn’t start with weeks of searching, benchmarking and optimization. Developers need to find the right model and understand its performance, while agents need clear signals to discover the same models, tools and information.
Today Arm is launching Arm AI Portal, giving developers and agents a common way to discover, optimize and deploy AI software across Arm compute. Developers can find task-specific, pre-optimized models with performance and accuracy data, compare latency, memory and size, and access code examples and deployment workflows.
AI Portal will soon provide tooling for developers to bring their own models, including proprietary models, for performance analysis and optimization on Arm. Agent-ready AI resources are available
Start faster with AI optimized for Arm
AI Portal supports language, speech, vision and neural graphics across Arm-based compute. At launch, pre-optimized models include Alibaba Qwen, Google Gemma and Ultralytics YOLO using runtimes including ExecuTorch, LiteRT and ONNX-RT, with support from ecosystem partners including Alibaba, Raspberry Pi and Ultralytics.
AI Portal meets developers and agents where they build, with Arm-optimized models available through Hugging Face and Portal resources accessible to coding agents through MCP.
Arm-optimized models are already delivering significant performance gains:
- Qwen3-TTS achieved an over 4x speedup on a vivo X300 smartphone using single-thread execution and mixed quantization with a Q8_0 talker and a code predictor, accelerated by Scalable Matrix Extension-2 (SME2).
- Ultralytics YOLO26n achieved over 40% performance improvement using single-thread execution with FP16 versus FP32 on a vivo X300 smartphone with SME2, and with FP16 and INT8 mixed quantization versus FP32 on Raspberry Pi 5 with NEON.
Enabling optimized AI across the Arm compute platform
AI Portalspans the Arm compute platform across cloud, edge and physical AI, helping developers identify models optimized for their target, from vision models for robotics to generative AI on smartphones to task-specific LLMs on cloud CPU.
AI Portal connects Arm technologies such as SVE, SME and neural acceleration with optimized software. Arm CSS for Mobile 2 is one example, with models accelerated by SME2 and GPUs with neural accelerators available through AI Portal.
For decades, Arm has invested in the software ecosystem. AI Portal extends that investment into the AI era, helping more than 22 million developers and their agents find and use software optimized for their target hardware.
Any re-use permitted for informational and non-commercial or personal use only.
Media Contacts
Kelly Tenn
Tech & Product Comms Manager, Developer
kelly.tenn@arm.com
+1 408 791 8467
Media Information
Media Contacts
Latest on X
AI is no longer just generating content. It’s completing tasks.
Across cloud, edge and physical systems, intelligence is becoming more capable — and compute has to keep up.
Arm is building the platform for that era.
We’ll share more from our Arm Everywhere China keynote,
Agentic AI is reshaping compute everywhere.
Across cloud, edge and physical AI, systems can increasingly perceive, reason and act — creating new demands for performance and efficiency at every layer.
See how Arm is meeting them. Discover more from Arm Everywhere China,
Agentic AI is changing what the data center needs from the CPU.
At SEMICON Taiwan, CK Tseng shared how Arm is helping the ecosystem scale AI infrastructure through more efficient CPUs, custom silicon, chiplets and tighter system integration.
As AI runs on-device, mobile platforms need to balance greater performance with efficiency.
Arm’s focus on performance per watt helps our partners deliver more capable mobile AI experiences, with mobile at the heart of a much bigger story across the edge.
Stay tuned for what’s
AI is changing how silicon gets designed.
On @NoPriorsPod, Rene Haas joins @Saranormous and @EladGil to discuss how AI could accelerate chip development, Arm’s evolution from IP to silicon, and why CPUs remain at the heart of computing as AI scales.
Listen
Developers should have the freedom to use the models that best fit what they’re building. Arm provides the compute foundation behind that flexibility. We’re proud to support the @NVIDIA and @HuggingFace partnership and the innovation it will unlock across the AI ecosystem.
Last week at World Robotics Conference, Federico Pecora shared how the next phase of robotics is becoming an increasingly complex systems challenge.
At Arm, we’re exploring the compute foundation needed to bring sensing, AI, planning, control and safety together reliably as
