1. Home
  2. Companies
  3. FuriosaAI
FuriosaAI logoFU

FuriosaAI

About

FuriosaAI, founded in 2017, designs AI chips built for efficiently running advanced models. The Seoul-based company specialises in AI chip design, Tensor Contraction Processor architecture, and deep learning hardware acceleration, with a focus on sustainable AI computing. Its team includes engineers with backgrounds at Samsung, AMD, and other leading semiconductor firms.

The company's product lineup spans hardware and software. Its second-generation RNGD accelerator targets high-performance large language model and multimodal inference, operating at just 180W and currently sampling with global technology partners. A first-generation Vision NPU has shipped into production. FuriosaAI also provides a software stack comprising a compiler, runtime, and integrations with PyTorch and Kubernetes.

Similar companies

Eridu AI logoEA

Eridu AI

Silicon Valley hardware startup building infrastructure solutions to accelerate AI model training and inference in data centres.

1 job
Fractile logoFR

Fractile

Fractile is a UK-based semiconductor company building AI acceleration hardware to radically improve frontier model inference performance.

Cerebras Systems logoCS

Cerebras Systems

Cerebras Systems designs wafer-scale AI chips and supercomputers for high-speed machine learning training and inference, serving enterprises, national labs, and healthcare systems.

femtoAI logoFE

femtoAI

femtoAI builds ultra-efficient AI inference technology and custom sparse processing hardware that enables artificial intelligence to run directly on edge devices without cloud connectivity.

quadric, Inc logoQI

quadric, Inc

Quadric develops general-purpose neural processing unit (GPNPU) IP and software tools that enable unified AI inference on edge devices, combining machine learning acceleration with full C++ programmability in a single processor architecture.

d-Matrix logoD-

d-Matrix

d-Matrix designs purpose-built AI inference computing platforms, using digital in-memory compute technology to run generative AI at scale efficiently and cost-effectively.