Compute & AI

Garmo-NPU-256

AI Compute — 256 tensor-slice edge NPU

Evaluation package
Request evaluation
GarCompute & AI
THE TECHNICAL ROLE

A clear integration boundary.

Garmo-NPU-256 integration role: AI Compute — 256 tensor-slice edge NPU. Documented composition: 256 tensor slices; each slice executes a 256-element INT4 dot product per cycle; INT2/INT4/INT8, FP4 E2M1/E3M0, FP8 E4M3/E5M2, BF16 accumulation; Native 2:4 and programmable N:M sparsity; zero-skip before operand fetch.

01

256 tensor slices

02

each slice executes a 256-element INT4 dot product per cycle

03

INT2/INT4/INT8, FP4 E2M1/E3M0, FP8 E4M3/E5M2, BF16 accumulation

04

Native 2:4 and programmable N:M sparsity

05

zero-skip before operand fetch

TECHNICAL OVERVIEW

The details matter.

Download the product brief
Canonical identity
IP-2066
Responsibility
AI Compute — 256 tensor-slice edge NPU
Documented interface architecture
AXI4 512-bit data master, AXI4-Lite CSR, 512-bit ready/valid tensor stream, IRQ/RAS/DFT
Documented building blocks
256 tensor slices; each slice executes a 256-element INT4 dot product per cycle; INT2/INT4/INT8, FP4 E2M1/E3M0, FP8 E4M3/E5M2, BF16 accumulation; Native 2:4 and programmable N:M sparsity; zero-skip before operand fetch
Delivery boundary
Evaluation / IP scope
Evidence scope
Evaluation package
Architecture targets · not measurements
196.6 TOPS INT4 dense @ 1.5 GHz; 393 TOPS effective with 2:4 sparsity; target <10 W core power on advanced node.
Evaluation package

Know the release.
Know what it establishes.

Canonical technology and its architecture/integration role are recorded. The offered delivery is confirmed against the exact release before licensing.

Source basis: Bible v7.42 IP-2066. This is a controlled-portfolio summary, not a fresh execution of the chip qualification flow.

Release-specific qualification

No node-port, silicon-performance, standards certification or analog qualification is inherited from a related product.

Canonical record: P1 / REPRESENTATIVE ARCHITECTURE RTL — concrete integration specification and synthesizable RTL shell/representative datapath only; all frequency, throughput, power, area, jitter and standards-superiority figures remain architecture targets until the required evidence hierarchy closes.

A PATH THAT FITS THE PROJECT

Build with Garmo-NPU-256.

Technical scope, rights, support and release configuration are agreed before delivery.

evaluation

Evaluate an individual IP configuration

Inspect the named release and establish technical fit.

Discuss this scope
production

Named-design production license

Agree deployment or design rights, deliverables and support.

Discuss this scope
custom

Node binding or derivative integration

Define the customer-specific work and acceptance criteria.

Discuss this scope

Keep exploring.

Compute & AI

NPJ SolveTile v1.0.0

Reusable solver/bitset/sparse tile binding lane ALU, popcount/zero-skip, descriptor FIFO and deterministic transport to a reference model and SV source.

RTL + models
IP-4276 · Source / reference evidence
Compute & AI

Garmo RISC-V Processor IP

CPU pipeline; cache/memory interfaces; MMU/TLB options; interrupt/debug; FPU/vector options; compiler/software support; verification/formal collateral; hardened physical views.

Architecture IP
IP-001 · Source documented
Compute & AI

Garmo-NPU-1024

AI Compute — 1024 tensor-slice cloud/robotics NPU

Evaluation package
IP-2067 · Source documented

Your next big idea.
Let’s build it together.

Start with a chip, a software release or a single IP block.

CANONICAL EVIDENCE RECORD · IP-2066

Source documented

A concrete source or implementation record exists; the selected evidence does not establish complete qualification.

Source: Bible 7.42 VERIFIED V29 · Bible v7.42 IP-2066.

Record integrity and qualification boundary

Retained exact-release source reference; no new table locator was inferred.

Evidence ranking for evaluation prioritization; not a probability or certification.

How evidence priority is determined ↗