TOP
background

LattePanda Mu ULTRA LattePanda Mu ULTRA

A micro x86 compute module for ON-DEVICE AI

Massive AI Power

Massive AI Power

  • Up to
    115 TOPS
    Overall Peak(Int8)
  • 64 TOPS
    GPU
  • 47 TOPS
    NPU
  • 4 TOPS
    CPU
  • Up to
    16 GB 8533MT/s
    MoP LPDDR5X
  • 8 Cores
    CPU 4.8GHz MAX
  • Up to
    8 Xe
    Arc Graphics
benchmark background

Redefining on-device Benchmarks

Local LLM
YOLO
CPU
GPU

Decode (text generation) speed · tokens/s (higher is better)

  • Qwen3.5-2B-int4-ov
    55.1
  • Qwen3.5-4B-int4-ov
    27.8
  • Qwen3.5-9B-int4-ov
    18.1
  • gemma-4-E2B-it-int4-ov
    36.0
  • gemma-4-E4B-it-int4-ov
    21.7
  • Phi-4-mini-instruct-int4-ov
    30.1
051015202530354045505560
Lightning-Fast Unified Memory On Package background

Lightning-Fast Unified Memory On Package

Up to 16GB LPDDR5X-8533RAM 136.5 GB/s high bandwidth for fast token throughput
Up to 11.6GBAllocatable VRAM for larger LLMs and KV cache
  • LattePanda Mu Ultra MoP 8533 MT/s
  • AMD Ryzen Al 9 365 Onboard 8000 MT/s
  • LattePanda Sigma (Intel Core i5-1340P) Onboard 6400 MT/s
Active or Idle, Maximize Battery Runtime background

Active or Idle, Maximize Battery Runtime

Active Mode

Higher Tokens. Lower Watts.

Perfect for fast AI Inference with maximum battery life.

Power Consumption(W) Decode Speed(tok/s)
  • 35W LattePanda Sigma (16GB) 17 tok/s
  • 26W LattePanda Mu Ultra (226V) 28 tok/s
Idle Mode

2.5W Ultra-Low Idle

Perfect for 24/7 always-on AI or battery-powered deployment

  • LattePanda Mu Ultra 2.5W
  • LattePanda Mu 4W

Card-Sized, Fits Somewhere

card-size
AI Software background

Works with Your AI Software Frameworks

Run LLMs, vision, and speech models using the tools you already know.

logos

Flexible I/O Expansion

  • Up to 4 PCIe 4.0 x1 Lanes
  • Up to 0 PCIe 4.0 x2 Lanes
  • Up to 0 PCIe 4.0 x4 Lanes
  • PCle 4.0 Lanes #8
  • PCle 4.0 Lanes #7
  • PCle 4.0 Lanes #6
  • PCle 4.0 Lanes #5
  • PCle 4.0 Lanes #4
  • PCle 4.0 Lanes #3
  • PCle 4.0 Lanes #2
  • PCle 4.0 Lanes #1
  • 2

    USB 3.2 Gen2
  • 3

    HDMI or DisplayPort
  • 6

    USB 2.0
  • 3

    UART
  • 3

    I2C
  • 14

    GPIO
  • 1

    CNVio3
  • Application Scenarios

    Smart AI Terminals

    Power offline LLMs for localized translation, voice assistants, and secure data queries without cloud dependency.

    Offline LLMPrivacy FirstVoice UI
    Smart AI Terminals

    Portable Instruments

    Desktop-grade x86 performance packed into handheld spectrum analyzers, diagnostic tools.

    Credit-Card SizeDesktop PowerCustom I/O
    Portable Instruments

    Autonomous Mobile Robots (AMR)

    Accelerate real-time sensor fusion and SLAM navigation while keeping payload weight to an absolute minimum.

    Sensor FusionROS2Light Payload
    Autonomous Mobile Robots (AMR)

    Service Robots

    Process multi-modal workloads (vision + voice) simultaneously in a compact "brain" designed robotics.

    Multi-modal AIReal-Time Control
    Service Robots

    Edge AI Vision

    Embed high-resolution video processing directly onto cameras or inspection arms for low-latency defect detection.

    Machine VisionLow LatencyLocal Analytics
    Edge AI Vision

    From ready-to-use hardware to full development support

    • Lite Carrier(V3)

      All essential I/O integrated into a standard 3.5-inch form factor. Ideal for rapid industrial deployment.

    • Mini Carrier

      Small yet highly integrated, featuring dual USB-C ports. Perfect for compact devices.

    • AI Accelerator Carrier

      Multiple M.2 slots for dedicated M.2 AI accelerators. Super scale your local AI compute limitlessly.

    Lite Carrier(V3)Mini CarrierAI Accelerator Carrier

    Resources & Support

    Specification

    Product LattePanda Mu Ultra 226V LattePanda Mu Ultra 256V
    Processor Intel® Core™ Ultra 5 Processor 226V Intel® Core™ Ultra 7 Processor 256V
    CPU 8 Cores, 8 Threads
    Up to 4.5 GHz
    8 Cores, 8 Threads
    Up to 4.8 GHz
    Memory 16GB LPDDR5X 8533 MT/s
    GPU Intel® Arc™ 130V GPU
    7 Xe-cores, Up to 1.85 GHz
    Intel® Arc™ 140V GPU
    8 Xe-cores, Up to 1.95 GHz
    NPU Intel® AI Boost 40 TOPS (Int8) Intel® AI Boost 47 TOPS (Int8)
    Overall Peak TOPS (Int8) 97 115
    Operating System Windows 11, Ubuntu 24.04
    Expansion Up to 4 PCIe 4.0 x1 Lanes
    Up to 4 PCIe 4.0 x2 Lanes
    Up to 2 PCIe 4.0 x4 Lanes
    2x USB 3.2 Gen2
    6x USB 2.0
    1x CNVio3
    3x UART
    3x I2C
    14x GPIO
    Display 3x HDMI or DisplayPort
    1x eDP
    Supports up to 3 independent displays simultaneously.
    Dimension 60mm * 69.6mm

    LattePanda Mu Ultra is a compact x86 AI compute module designed for on-device AI, edge computing, and embedded AI systems. It is a high-performance computing platform based on Intel Core Ultra processors, combining desktop-class x86 computing with integrated AI acceleration in a small form factor. Mu Ultra is built for developers, engineers, and OEMs who need powerful local AI processing, flexible hardware integration, and a scalable compute solution for next-generation intelligent devices.

    LattePanda Mu Ultra features Intel Core Ultra 5 and Core Ultra 7 processors with up to 8 CPU cores, Intel Arc Graphics with up to 8 Xe-cores, and Intel AI Boost NPU delivering up to 48 TOPS AI acceleration. The platform provides up to 115 TOPS INT8 AI performance combining CPU, GPU, and NPU, up to 32GB LPDDR5X-8533 high-speed memory, up to 24GB GPU-allocatable memory, and a compact 69.6 × 60mm module design. It supports PCIe 4.0 expansion, multiple USB interfaces, GPIO, UART, I2C, CNVio connectivity, and up to three independent displays. With support for Windows and Linux, as well as AI frameworks including Intel OpenVINO, llama.cpp, and Ollama, Mu Ultra provides a flexible foundation for local LLM inference and AI application development.

    LattePanda Mu Ultra enables a wide range of AI-powered applications, including local AI agents, AI home servers, robotics, autonomous machines, computer vision systems, smart terminals, industrial automation, edge AI gateways, and intelligent embedded devices. By bringing x86 compatibility, AI acceleration, and rich hardware expansion into a compact compute module, Mu Ultra helps developers move AI workloads from cloud platforms to real-world products. It is an ideal solution for edge AI developers, makers, system integrators, and OEM companies building the next generation of AI devices.

    TOP