Arm Newsroom News
News

Arm unveils Arm AI Portal to accelerate optimized AI apps across the Arm compute platform 

By Sharbani Roy, VP, AI & Developer Platforms, Arm

News highlights

  • Arm AI Portal connects more than 22 million developers and their agents to optimized AI software across the Arm compute platform spanning cloud, edge and physical AI  
  • Developers can start faster with pre-optimized models or bring and optimize their own models for best performance on Arm 
  • As development becomes increasingly agentic, AI Portal makes models, performance data and workflows machine-discoverable, meeting developers and agents where they already build, giving simpler access to Arm’s latest AI technologies  

Agentic AI is moving beyond the cloud to edge and physical AI, creating complexity as developers build across models, runtimes and hardware targets. Arm’s compute platform spans this continuum, supported by more than 22 million developers who need optimized AI software. 

Building an AI application shouldn’t start with weeks of searching, benchmarking and optimization. Developers need to find the right model and understand its performance, while agents need clear signals to discover the same models, tools and information. 

Today Arm is launching Arm AI Portal, giving developers and agents a common way to discover, optimize and deploy AI software across Arm compute. Developers can find task-specific, pre-optimized models with performance and accuracy data, compare latency, memory and size, and access code examples and deployment workflows.  

AI Portal will soon provide tooling for developers to bring their own models, including proprietary models, for performance analysis and optimization on Arm. Agent-ready AI resources are available via early access, ahead of general release. 

Start faster with AI optimized for Arm 

AI Portal supports language, speech, vision and neural graphics across Arm-based compute. At launch, pre-optimized models include Alibaba Qwen, Google Gemma and Ultralytics YOLO using runtimes including ExecuTorch, LiteRT and ONNX-RT, with support from ecosystem partners including Alibaba, Raspberry Pi and Ultralytics. 

AI Portal meets developers and agents where they build, with Arm-optimized models available through Hugging Face and Portal resources accessible to coding agents through MCP. 

Arm-optimized models are already delivering significant performance gains:  

  • Qwen3-TTS achieved an over 4x speedup on a vivo X300 smartphone using single-thread execution and mixed quantization with a Q8_0 talker and a code predictor, accelerated by Scalable Matrix Extension-2 (SME2)
  • Ultralytics YOLO26n achieved over 40% performance improvement using single-thread execution with FP16 versus FP32 on a vivo X300 smartphone with SME2, and with FP16 and INT8 mixed quantization versus FP32 on Raspberry Pi 5 with NEON. 

Enabling optimized AI across the Arm compute platform

AI Portal spans the Arm compute platform across cloud, edge and physical AI, helping developers identify models optimized for their target, from vision models for robotics to generative AI on smartphones to task-specific LLMs on cloud CPU. 

AI Portal connects Arm technologies such as SVE, SME and neural acceleration with optimized software. Arm CSS for Mobile 2 is one example, with models accelerated by SME2 and GPUs with neural accelerators available through AI Portal. 

For decades, Arm has invested in the software ecosystem. AI Portal extends that investment into the AI era, helping more than 22 million developers and their agents find and use software optimized for their target hardware. 

Build AI on Arm, faster

Find optimized models, deployment-ready code, workflows, and optimization tools to build and deploy AI applications on Arm, from cloud to edge.

Article Text
Copy Text

Any re-use permitted for informational and non-commercial or personal use only.

Media Contacts

Kelly Tenn
Tech & Product Comms Manager, Developer
+1 408 791 8467
Sign Up for Media & Analyst News
Get the latest media & analyst news direct from Arm

Latest on X

promopromopromopromopromopromopromopromo