Bertha 500

Bertha 500 is HyperAccel’s flagship LPU-based AI inference accelerator, purpose-built for large language models. Unlike general-purpose GPUs, its domain-specific architecture maximizes memory bandwidth utilization, delivering high-speed token generation at a fraction of the power and cost. Bertha 500 supports popular open LLMs and enables enterprises to run chatbots, document AI, and generative AI services securely on their own infrastructure. Available as PCIe accelerator cards and integrated server or workstation systems, it scales from single-node deployment to data-center clusters. Live demo at our booth: real-time LLM inference running on Bertha 500 hardware.

en_USEN