1. Home
  2. Companies
  3. Fractile
Fractile logoFR

Fractile

About

We're building a new kind of computing architecture designed specifically for the AI inference bottleneck. Our approach integrates memory and compute in a way that eliminates the data movement overhead that limits traditional GPUs. The team spans from transistor-level circuit design up to cloud inference server logic - we work across the full stack because that's the only way to solve this problem properly.

Founded in 2022 and operating out of London and Bristol, we emerged from stealth in 2024 with backing from NATO's Innovation Fund, Oxford Science Enterprises, and Kindred Capital. Our goal is straightforward: run frontier AI models 25x faster at 1/10th the cost. We're taking a hardware-software co-design approach, building our own software stack alongside custom silicon rather than trying to patch over the limitations of general-purpose hardware.

Similar companies

Eridu AI logoEA

Eridu AI

Silicon Valley hardware startup building infrastructure solutions to accelerate AI model training and inference in data centres.

1 job
FuriosaAI logoFU

FuriosaAI

FuriosaAI designs AI inference chips and a supporting software stack for efficient, high-performance model execution, with products shipping to global technology partners.

Cerebras Systems logoCS

Cerebras Systems

Cerebras Systems designs wafer-scale AI chips and supercomputers for high-speed machine learning training and inference, serving enterprises, national labs, and healthcare systems.

d-Matrix logoD-

d-Matrix

d-Matrix designs purpose-built AI inference computing platforms, using digital in-memory compute technology to run generative AI at scale efficiently and cost-effectively.

Sciforium logoSC

Sciforium

Sciforium builds AI infrastructure, including multimodal foundation models and a high-efficiency serving platform for rapid model deployment.

Inference logoIN

Inference

Inference.net runs a distributed GPU cluster providing AI inference infrastructure, custom model training, and deployment services designed to reduce costs and improve speed for production AI workloads.