Nota AI has spent years proving that smaller AI models can still punch above their weight. Now, the Korean AI optimization specialist is taking that philosophy squarely into high-performance computing.
The company announced it has signed a technology supply contract with FuriosaAI to provide its AI model optimization platform, NetsPresso®, for FuriosaAI’s flagship neural processing unit (NPU), RNGD—pronounced “Renegade.” The deal marks a notable shift for Nota AI, whose technology has traditionally been associated with edge devices like mobile phones, PCs, and automotive systems.
This time, the target is far more ambitious: AI servers and data centers running large-scale models at industrial speed.
From Edge Efficiency to Data Center Scale
NetsPresso® is designed to compress and optimize AI models—cutting model size by as much as 90% while preserving accuracy. Until recently, that capability was most often framed as an edge AI advantage: enabling inference on devices with limited compute, power, or memory.
The FuriosaAI contract reframes that narrative. By optimizing models for the RNGD NPU, Nota AI is positioning NetsPresso® as a critical layer not just for constrained devices, but for high-performance AI infrastructure, where efficiency translates directly into lower costs, higher throughput, and more predictable performance.
In practical terms, Nota AI’s optimization technology will be used to improve inference speed and stability for large AI models running on RNGD. For customers deploying AI at scale—whether in factories, logistics centers, or data-intensive industrial environments—that combination could mean faster responses, lower power draw, and more efficient utilization of hardware.
Why RNGD Matters
FuriosaAI has been steadily building its reputation as a serious contender in the custom AI silicon market, particularly as enterprises look for alternatives to GPU-centric inference stacks. RNGD is designed to deliver high performance and efficiency for AI workloads, especially in scenarios where predictable latency and power efficiency matter more than brute-force training performance.
By pairing RNGD with NetsPresso®, FuriosaAI gains a software advantage that many hardware vendors struggle to deliver: deep, hardware-aware model optimization. Rather than forcing customers to adapt generic models to specialized silicon, the partnership promises AI models that are tuned from the outset to run efficiently on FuriosaAI’s architecture.
That matters in a market increasingly crowded with NPUs, accelerators, and domain-specific chips all competing on performance-per-watt and total cost of ownership.
A Strategic Expansion for Nota AI
For Nota AI, the deal represents more than a single supply contract. It signals an expansion into the AI server and data center semiconductor ecosystem, a space dominated by hyperscalers, GPU vendors, and increasingly, custom silicon startups.
Until now, Nota AI’s customer base has centered on mobile, PC, and automotive application processors. With FuriosaAI, the company extends its reach across the full AI deployment spectrum—from edge to core.
That breadth could prove strategically important. As enterprises deploy AI across heterogeneous environments—cameras at the edge, robots on factory floors, and centralized inference in data centers—the ability to optimize models consistently across hardware becomes a competitive differentiator.
Nota AI’s leadership is clearly leaning into that vision. CEO Myungsu Chae described the contract as validation that NetsPresso®’s optimization technology has commercial value well beyond on-device AI, reinforcing its role in high-performance data center environments.
Beyond a Supply Deal: A Joint Business Model
The announcement goes further than silicon and software integration. Nota AI and FuriosaAI have also agreed to establish a strategic partnership built around Nota AI’s vision AI solution, Nota Vision Agent (NVA).
By packaging NVA with FuriosaAI’s RNGD NPU, the companies are moving toward a turnkey offering that combines optimized hardware, compressed models, and application-ready vision AI. This bundled approach lowers the barrier for customers who want deployable AI systems without stitching together components from multiple vendors.
The partnership follows a technical cooperation memorandum of understanding signed in November, but the latest agreement formalizes a path from joint R&D to commercialization. In other words, this is no longer just experimentation—it’s a coordinated go-to-market strategy aimed at accelerating revenue on both sides.
Implications for the AI Hardware Market
The collaboration highlights a broader trend in the AI industry: hardware alone is no longer enough. As AI models grow larger and deployment environments more diverse, optimization software is becoming as critical as the silicon itself.
GPU vendors, NPU startups, and accelerator designers are all grappling with the same challenge—how to deliver consistent performance and efficiency across real-world workloads. Model compression, pruning, and hardware-aware optimization are increasingly central to that equation.
By aligning early and deeply, Nota AI and FuriosaAI are betting that customers will favor integrated stacks over piecemeal solutions. It’s a strategy reminiscent of how NVIDIA pairs CUDA and software libraries with its GPUs, albeit focused on inference efficiency rather than training dominance.
Industrial and “Physical AI” Opportunities
Both companies are explicit about where they see growth. FuriosaAI expects the partnership to increase the competitiveness of its NPU products across a wide range of industrial sites. Nota AI, meanwhile, sees new opportunities in sectors such as physical AI—systems that combine perception, reasoning, and action in the real world.
Vision-driven AI applications, from smart factories to autonomous systems, demand fast, reliable inference under tight power and latency constraints. A packaged NPU-plus-optimized-model solution directly addresses those requirements.
If successful, the approach could appeal to enterprises looking to deploy AI without relying exclusively on cloud-based inference or power-hungry GPUs.
A Showcase for Korea’s AI Ecosystem
There’s also a national dimension to the announcement. Both CEOs emphasized the partnership as a demonstration of Korea’s growing competitiveness in AI hardware and software innovation.
In a global market dominated by U.S. and Chinese tech giants, collaborations like this signal an effort to build vertically integrated AI stacks capable of competing internationally—particularly in industrial and enterprise deployments.
Whether the partnership gains traction beyond early adopters will depend on execution. But the strategic intent is clear: leaner models, specialized silicon, and tightly integrated solutions designed for the realities of AI at scale.
Power Tomorrow’s Intelligence — Build It with TechEdgeAI









