Sign in

FuriosaAI

@furiosa.ai
32 followers 0 following 97 posts

Our mission is to make AI computing sustainable, enabling everyone on Earth access to powerful AI.

PostsRepliesMedia
FuriosaAI @furiosa.ai · 13/02/2026
Thank you to the leaders and experts who joined us for a productive discussion on advancing the industry together and developing better ways to meet the world’s demand for AI compute.
010
FuriosaAI @furiosa.ai · 13/02/2026
During the visit, we shared a live demo of state-of-the-art LLMs running on Furiosa’s RNGD accelerator, which is now in mass production. It was a great opportunity to showcase how Furiosa is delivering high-performance, energy-efficient inference that moves beyond the limitations of GPUs.
110
FuriosaAI @furiosa.ai · 13/02/2026
We welcomed South Korea’s Minister of Trade, Industry and Resources, Jung-Kwan Kim, and other industry leaders to FuriosaAI HQ to discuss the future of the AI semiconductor ecosystem.
210
FuriosaAI @furiosa.ai · 29/01/2026
RNGD enters mass production. The market has been demanding a high-performance alternative to existing general-purpose and matrix-multiplication-based architectures that runs efficiently in standard data centers. Until now, there hasn’t been a strong option. Read the announcement: lnkd.in/gMfGg-ze
000
FuriosaAI @furiosa.ai · 30/12/2025
Looking ahead, while the industry will continue to evolve in ways no one can fully predict, our focus is clear: making inference seamless, energy-efficient, and production-ready at scale. The next phase of AI inference is here. More to come from Furiosa in 2026 🚀
000
FuriosaAI @furiosa.ai · 30/12/2025
This year’s progress includes: → Real-world validation, delivering up to 2.25× better performance per watt compared to GPUs → Major SDK advances across multi-chip scaling and tensor parallelism → Introduction of the NXT RNGD Server → A $125M Series C bridge to accelerate our next phase of execution
100
FuriosaAI @furiosa.ai · 30/12/2025
Over the past year, we’ve moved from architectural conviction to deploying inference systems built for what’s next. Working with partners like LG AI Research and OpenAI, we’ve validated a core belief: the future of AI acceleration starts with rethinking the fundamental compute primitive.
100
FuriosaAI @furiosa.ai · 30/12/2025
Inference has quietly become one of the largest cost centers behind the adoption of agentic AI. As models scale and workloads move from experimentation to production, the economics of inference are forcing organizations to rethink infrastructure to keep AI practical, reliable, and sustainable.
110
FuriosaAI @furiosa.ai · 30/12/2025
AI is moving fast. From single-model prompts to agentic, multi-model systems running continuously in production, we’ve entered a new inference reality.
100
FuriosaAI @furiosa.ai · 03/12/2025
Come visit our poster and learn more about our commitment to deep research with real-world impact 🌎 Read the paper here: arxiv.org/abs/2507.06996
000
FuriosaAI @furiosa.ai · 03/12/2025
Exciting research news from FuriosaAI + KAIST! Our collaborative project, RawMed, has been accepted at #NeurIPS2025. This pioneering framework synthesizes multi-table, time-series clinical data using text-based representations, compression techniques, and efficient autoregressive modeling.
100
FuriosaAI @furiosa.ai · 01/12/2025
We’re excited to carry this momentum into 2026 as RNGD moves toward broader availability. www.crn.com/news/compone...
000
FuriosaAI @furiosa.ai · 01/12/2025
We are honored to be named one of CRN's 10 Hottest Semiconductor Startups of 2025. This recognition highlights three recent critical milestones: our successful Series C funding, major customer validation from LG AI Research, and our first RNGD server product.
100
FuriosaAI @furiosa.ai · 25/11/2025
We're headed to #NeurIPS2025! We’re excited to join the global AI community in San Diego next week to share research insights, connect with innovators, and explore the latest breakthroughs in machine learning. 📍 Stop by our booth S13 to meet the team and discover opportunities to collaborate.
010
FuriosaAI @furiosa.ai · 21/11/2025
🎉 What an incredible evening co-hosting The AI Frontier with Akamai! It was inspiring to bring together leaders, experts, and researchers across industries to share real-world use cases, trends, and challenges in building scalable and sustainable AI.
020
FuriosaAI @furiosa.ai · 12/11/2025
In an interview with Bloomberg, our CEO June Paik discusses Furiosa’s growing traction with enterprise customers and why the future of AI depends on energy-efficient infrastructure, not just raw compute power. ⚡ Watch the full interview from #AISummitSeoul: www.bloomberg.com/news/videos/...
bloomberg.com
FuriosaAI CEO on Business Strategy, AI Policy
June Paik, founder and CEO at chip startup FuriosaAI, discusses the company's business strategy and South Korea's push for AI investment and policy support. He speaks with Shery Ahn from the sidelines...
010
FuriosaAI @furiosa.ai · 07/11/2025
Learn more and register here: luma.com/hes7s3du?tk=...
000
FuriosaAI @furiosa.ai · 07/11/2025
Join us for an evening of conversation and networking with leaders shaping the future of AI. We are excited to co-host this in-person event with Akamai Technologies, bringing together leaders from across industries and borders to explore how to design and scale high-performance AI systems.
100
FuriosaAI @furiosa.ai · 31/10/2025
To showcase this, we're powering a live multilingual demo, representing all 21 APEC member economies, on OpenAI's gpt-oss-120b model running efficiently on just two RNGD chips. Interested in the future of efficient compute? Come talk to our team to learn more.
000
FuriosaAI @furiosa.ai · 31/10/2025
The future of AI runs on efficient compute. At #APEC2025, we’re demonstrating what that looks like 🚀 Our RNGD chip was engineered from the ground up with AI-native architecture to redefine compute, delivering uncompromising performance with breakthrough power efficiency.
100
FuriosaAI @furiosa.ai · 22/10/2025
📌 We’re at #PyTorchCon this week! As a global company, we are excited to see that 61 countries contributed to @pytorch.org this year, not to mention all the vLLM, DeepSpeed, and Ray updates. 🤝 Connect with us at Booth S13 on October 22 and October 23 to discuss your AI infrastructure stack.
000
FuriosaAI @furiosa.ai · 17/10/2025
Read the story here 👉https://furiosa.ai/blog/serving-gpt-oss-120b-at-5-8-ms-tpot-with-two-rngd-cards-compiler-optimizations-in-practice
000
FuriosaAI @furiosa.ai · 17/10/2025
Behind the scenes: Furiosa engineer Sanguk Park and CTO Hanjoon Kim share how our team served gpt-oss-120B at 5.8 ms TPOT using just two RNGD cards while maintaining exceptional power efficiency (under 180 W per card) for OpenAI's South Korean office grand opening.
100
FuriosaAI @furiosa.ai · 15/10/2025
Register here: luma.com/dwd30txb
010
FuriosaAI @furiosa.ai · 15/10/2025
The 2025 OCP Global Summit is officially underway 🎉 Join FuriosaAI, hosted·ai, and BlueSky Compute for a happy hour. Connect with peers across the AI infrastructure, cloud, and systems communities over drinks and light bites. Meet the team and exchange insights on next-generation AI compute.
110
FuriosaAI @furiosa.ai · 10/10/2025
Performance gains vs. 2025.2.0: ✅ Llama 3.1 8B: Up to 4.5% average throughput boost and up to 55% average reduction in Time-to-First-Token (TTFT) ✅ Llama 3.1 70B: Up to 3× average throughput improvement and up to 35% TTFT reduction
000
FuriosaAI @furiosa.ai · 10/10/2025
Missed our latest SDK 2025.3.0 release? We’re making #furiousprogress unlocking new levels of performance and efficiency for large-scale models and agentic AI on RNGD. 📑 More details: furiosa.ai/blog/furiosa... 📋 Release notes: developer.furiosa.ai/latest/en/wh...
100
FuriosaAI @furiosa.ai · 10/10/2025
It’s been an exciting few days at #SDEC2025 ⚡ Thank you to everyone who visited our booth and joined our keynote and panel on Furiosa’s journey to building next-gen AI chips. Drop by Booth 6307 today and tomorrow as the conference continues to meet the team and keep the conversation going.
010
FuriosaAI @furiosa.ai · 07/10/2025
Read the full announcement here: www.bytebt.com/bytebridge-s...
000
FuriosaAI @furiosa.ai · 07/10/2025
"Furiosa is committed to delivering cutting-edge processors that power the next wave of AI workloads. By working with ByteBridge, we can ensure that these innovations are deployed at scale, supporting resilient digital infrastructure across the region." – Alex Liu, SVP of Product and Business
100
FuriosaAI @furiosa.ai · 07/10/2025
We’re pleased to announce a new partnership between FuriosaAI and ByteBridge to advance next-generation AI infrastructure across the APAC region.
100
FuriosaAI @furiosa.ai · 02/10/2025
🤝 Visit us at KLCC Hall 6, Booth 6037. We’d love to connect with you.
000
FuriosaAI @furiosa.ai · 02/10/2025
🛬 We’re headed to #SDEC2025 in Kuala Lumpur next week to explore the future of AI and semiconductors. 👀 Don’t miss our co-founder and CEO, June Paik, onstage October 10 as he shares how Furiosa is building next-gen AI chips for a more sustainable future.
100
FuriosaAI @furiosa.ai · 30/09/2025
Under the hood: 🔹 Dual CPUs + optimized PCIe topology tuned for RNGD 🔹 Enterprise-grade reliability with dual management paths + secure boot 🔹 Cloud-native from day one with pre-installed SDK and vLLM-compatible serving interface 👉 Learn more: furiosa.ai/blog/introdu...
000
FuriosaAI @furiosa.ai · 30/09/2025
The result: dramatically lower power costs, higher deployment density, and the ability to build AI Factories with faster token throughput. Ready to serve LLMs from day one — no integration delays, no specialized power or cooling required.
100
FuriosaAI @furiosa.ai · 30/09/2025
Introducing Furiosa NXT RNGD Server 🚀 Engineered for efficient AI inference, the NXT RNGD Server hosts up to 8 RNGD accelerators, delivering: ⚡ 4 PFLOPS (FP8) compute 📦 384GB HBM + 12TB/s bandwidth 🔋 Just 3kW power draw (compared to 10.2kW for H100 SXM servers)
100
FuriosaAI @furiosa.ai · 17/09/2025
Last week, we demoed OpenAI’s open-weight gpt-oss 120B model running live on RNGD, our flagship AI accelerator. Here’s a recording of the demo in action, showing that cutting-edge models can be deployed well within the existing power budgets of typical data centers:
000
FuriosaAI @furiosa.ai · 11/09/2025
This collaboration with OpenAI highlights the essential role that AI-native hardware like RNGD plays in building more economically and environmentally sustainable AI for enterprises around the world.
000
FuriosaAI @furiosa.ai · 11/09/2025
The combination of RNGD's performance and power efficiency with advanced new models like gpt-oss will play a critical role in making advanced AI more sustainable and accessible.
100
FuriosaAI @furiosa.ai · 11/09/2025
This setup demonstrates that cutting-edge models can be deployed within typical data centers' existing power budgets, removing the prohibitive energy costs and complex infrastructure requirements of GPUs and making advanced AI truly accessible for enterprise customers.
100
FuriosaAI @furiosa.ai · 11/09/2025
Today we partnered with OpenAI for the grand opening of its new Seoul office, where we showcased its new open-source gpt-oss 120B model running live on RNGD. We demonstrated a real-time chatbot efficiently running the model on just two RNGD cards, using MXFP4 precision: furiosa.ai/blog/furiosa...
100
FuriosaAI @furiosa.ai · 10/09/2025
Thank you to everyone who has already stopped by to connect with us. If you haven’t yet, we’d love to meet and discuss our Tensor Contraction Processor (TCP) and our flagship AI accelerator, RNGD, at 📌 Booth #728!
000
FuriosaAI @furiosa.ai · 10/09/2025
Day 1 at #AIInfraSummit was energizing. The conversations at our booth make it clear: the future of AI depends not just on smarter models, but on more efficient compute. That’s why we were excited to have Alex Liu onstage sharing how we are redefining performance and efficiency in AI infrastructure.
100
FuriosaAI @furiosa.ai · 05/09/2025
🛬 Next week, we’re at #AIInfraSummit. 👀 Watch our SVP of Product and Business, Alex Liu, present on the Enterprise AI stage at 4:10 PM on September 9. 🤝 Connect with us in the app, stop by Booth 728 from September 9 to September 11, and book a meeting in advance: lp.furiosa.ai/ai-infra-sum...
000
FuriosaAI @furiosa.ai · 03/09/2025
Our engineers tackle groundbreaking challenges head-on. This month, we feature Seung Ho Song, a key engineer on our compiler team, and ask him about his experience. Read the full spotlight in our newsletter: www.linkedin.com/pulse/furios...
000
FuriosaAI @furiosa.ai · 27/08/2025
The annual Hot Chips conference this week reminded us to share our article from the recent special issue of IEEE Micro, which highlights our Hot Chips 2024 presentation as one of the best at the event. 👀 Read the full article here: dxttx52ei7rol.cloudfront.net/Micro_202503...
010
FuriosaAI @furiosa.ai · 21/08/2025
🔥 From our LG AI Research news to our $125 million funding raise, and from new executive hires to presenting six papers at ICML and ACL, it’s been a busy few months. We also attended events in Singapore, Paris, and NYC. Read our latest newsletter for details. 🗞️ www.linkedin.com/pulse/furios...
linkedin.com
FuriosaAI news: LG AI Research, $125M in funding, new executives, and more
LG AI Research taps FuriosaAI to achieve 2.25x better LLM inference performance vs.
000
FuriosaAI @furiosa.ai · 19/08/2025
🙏 Thank you to our team and all our collaborators, including Kangwook Lee, Yuchen Zeng, Thomas Zeng, Coleman Hooper, Amir Gholami, and Kurt Keutzer. 👀 Learn more: furiosa.ai/blog/furiosa...
furiosa.ai
Furiosa presents papers about improving AI performance at ICML 2025…
Last month, we presented four papers at ICML 2025 in Vancouver and two papers at ACL 2025 in Vienna, authored by our employees and our collaborators.
000
FuriosaAI @furiosa.ai · 19/08/2025
We’re grateful to have collaborated on these projects with the talented researchers at Korea University, Seoul National University, University of Wisconsin-Madison, Ajou University, UC Berkeley, UC San Francisco, ICSI, LBNL, Microsoft’s Gray Systems Lab, and University of Lisbon.
200
FuriosaAI @furiosa.ai · 19/08/2025
Last month, we presented four papers at ICML 2025 in Vancouver and two papers at ACL 2025 in Vienna. These six papers dig into ways to make advanced AI systems more efficient, more capable, and more flexible.
100