Phison aiDAPTIV+ is Cost-Effective With Large Models to Enable Generative AI

Phison aiDAPTIV+ is Cost-Effective With Large Models to Enable Generative AI

Analyst(s): Alastair Cooke
Publication Date: May 30, 2025

What is Covered in this Article:

  • Phison aiDAPTIV+ enables cost-effective running of large AI models on modest GPUs
  • The combination of the aiDAPTIVCache hardware and aiDAPTIVLink drivers runs unmodified PyTorch applications
  • Training is vital to the successful adoption of AI into an organization

The Event – Major Themes & Vendor Moves: AI Infrastructure Field Day is a semiannual, invitation-only event held in Santa Clara, organized by Tech Field Day. Independent industry experts join presenting companies to learn about product innovations. The event is live-streamed, and then videos are published on the Tech Field Day YouTube channel.

Phison aiDAPTIV+ is Cost-Effective With Large Models to Enable Generative AI

Analyst Take: Using a fast SSD as a tier to expand usable RAM is not a new concept; most operating systems have swap files to enable more system memory than the installed RAM. Phison aiDAPTIV+ is a set of hardware and software that enables the same concept for GPU memory, making AI cost-effective with large models. The high-bandwidth memory in a GPU is a significant part of the GPU’s price. As AI models become larger, the need for more memory follows. The fastest option is still to keep the entire model and data in GPU memory, which requires high-end GPUs for large models and often involves using cluster GPUs to accommodate the whole model. Phison aiDAPTIV+ is not intended to address use-cases that require the highest possible training throughput or the lowest inference latency, where the performance justifies the cost of the GPUs. Phison aiDAPTIV+ will deliver cost-effective results in use cases where lower throughput or longer response latency is acceptable for a significant cost saving. For example, for a 66% cost reduction, a weekly task of fine-tuning a model might be completed overnight, rather than in two hours. The aiDAPTIV+ solution comprises the aiDAPTIVcache hardware and aiDAPTIVlink drivers, with the aiDAPTIVProSuite as an optional AI software development environment.

aiDAPTIVcache

The SSDs Phison uses in AI acceleration have custom firmware designed to extend the SSD’s lifespan, exceeding that of the PC or embedded device where it is installed. Available in M.2 form at up to 320 GB for laptops and PCs, and in U.2 form at up to 8 TB for workstations and servers, the SSDs are part of the Pascari range, which Phison recently introduced directly to the market. The optimized firmware provides 100 DWPD (drive writes per day), allowing both high performance and high endurance for the most demanding generative AI workloads. These SSDs can be installed in embedded devices, such as the NVIDIA Jetson Nano, or laptops and desktops with desktop-class GPUs, making these devices cost-effective with large models

aiDAPTIVlink

The real magic lies in the aiDAPTIVlink software layer, which sits between the standard PyTorch library and the combination of SSD and GPU. The software manages moving blocks of data between GPU memory and the SSD as needed, enabling unmodified PyTorch applications to utilize larger models and more contextual data. Not requiring application changes makes aiDAPTIV+ a simple option to deploy compared to reducing model size through quantization or buying more expensive GPUs.

aiDAPTIVPro Suite

The whole aiDAPTIV+ solution originated from Phison’s internal challenges with adopting generative AI within the business. The aiDAPTIVPro Suite is a graphical tool for building LLM-based training and inference; it is an optional component in aiDAPTIV+. It reflects Phison’s understanding that training is crucial for enabling AI adoption. Phison has a training program for bringing staff up to speed with AI, delivered by select partners for their end-customers who will deploy aiDAPTIV+.

What to Watch:

  • Phison aiDAPTIV+ is not intended as a general solution for running large generative AI models, only for situations where the lower performance is justified by lower cost.
  • Other SSD vendors have demonstrated similar techniques, although not yet as integrated as the Phison solution. We expect to see more solutions that address the high cost of GPUs and the disparity between the price of GPU memory and SSD storage.
  • Moves towards smaller language models for specific purposes may reduce the need for this memory tiering with GPUs during inference. The need for more memory during training is unlikely to change significantly, as quantization to reduce model size typically occurs after training.

The Phison presentations at AI Infrastructure Field Day are available on their appearance page. You can watch all the presentations from the four days of AI Infrastructure Field Day on the Tech Field Day website.

Disclosure: Futurum is a research and advisory firm that engages or has engaged in research, analysis, and advisory services with many technology companies, including those mentioned in this article. The author does not hold any equity positions with any company mentioned in this article.

Analysis and opinions expressed herein are specific to the analyst individually and data and other information that might have been provided for validation, not those of Futurum as a whole.

Other insights from Futurum:

Organizations Face Triple Threat in Generative AI: Privacy, Talent & Compliance

DeepSeek Disrupts AI Market with Low-Cost Training and Open Source, Yet Many Questions Loom

Channel Partner Ecosystem is Ready to Capitalize on Generative AI

Author Information

Alastair has made a twenty-year career out of helping people understand complex IT infrastructure and how to build solutions that fulfil business needs. Much of his career has included teaching official training courses for vendors, including HPE, VMware, and AWS. Alastair has written hundreds of analyst articles and papers exploring products and topics around on-premises infrastructure and virtualization and getting the most out of public cloud and hybrid infrastructure. Alastair has also been involved in community-driven, practitioner-led education through the vBrownBag podcast and the vBrownBag TechTalks.

Related Insights
Meta Reopens Its Models. Is This a PC Play or a Policy Play?
August 10, 2026

Meta Reopens Its Models. Is This a PC Play or a Policy Play?

Nick Patience, VP and Practice Lead, AI Platforms at Futurum, examines Meta’s return to open weights with Muse Glimmer and what Mark Zuckerberg’s superintelligence letter reveals about the company’s real...
AMD Q2 FY 2026 EPYC and Helios Fuel the Next AI Growth Phase
August 7, 2026

AMD Q2 FY 2026: EPYC and Helios Fuel the Next AI Growth Phase

Brendan Burke, Research Director at Futurum, analyzes AMD’s Q2 FY 2026 earnings, focusing on Data Center growth, EPYC demand, Helios momentum, and the next phase of AI infrastructure execution....
Autonomous AI
August 5, 2026

Wayve and Uber’s Licensing Milestone: A Major shift for Autonomous Rides?

Transport for London grants vehicle licences to Wayve and Uber, marking a pivotal regulatory milestone for autonomous AI in public urban streets. Early rider access planned for summer 2026 as...
August 4, 2026

Uncrewed Vessels Set to Revolutionize Anti-Submarine Warfare for Dutch Navy

Thales's selection to design uncrewed vessels for the Royal Netherlands Navy underscores OT/ICS security as defense's greatest challenge, with quantum-safe cryptography now standard....
Ingram Micro Q2 FY 2026: Xvantage Strengthens the AI Distribution Model
August 3, 2026

Ingram Micro Q2 FY 2026: Xvantage Strengthens the AI Distribution Model

Futurum Research analyzes Ingram Micro’s Q2 FY 2026 earnings, focusing on AI infrastructure, cloud growth, Xvantage adoption, and channel consolidation....
Qualcomm Q3 FY 2026: Automotive Growth Offsets Handset Weakness
August 3, 2026

Qualcomm Q3 FY 2026: Automotive Growth Offsets Handset Weakness

Futurum Research analyzes Qualcomm’s Q3 FY 2026 earnings, focusing on handset pressure, automotive growth, and the data center ramp....

Book a Demo

Welcome

The vision behind everything in Futurum’s Custom Research practice is this: research should show you what is happening, what comes next, and what to do about it. It should be personal to each audience, easy for people to grasp, and structured so LLMs can reason over it accurately. And it should be fast and turnkey; you want answers now, not another project to carry for quarters.

Whether you are defining business, channel, or go-to-market strategy; evaluating vendors or justifying ROI; or commissioning research to fill an emerging market need, we have your back, with a program that answers your questions with the objectivity and credibility to drive real decisions.

To do it, we bring unmatched data to bear: Futurum research, surveys, and market projections; validated market feeds; ETR’s 15 years of insight from 10,000 technology decision-makers; G2’s buyer and user data; and what our analysts hear every day. Add leading primary collection, from AI-moderated voice interviews to surveys and analyst-led interviews, all turnkey, and every project comes out credible, nuanced, and actionable.

And we don’t just drop the results in your lap. For internal work, we provide analyst-led sessions, interactive dashboards, and a range of formats. For market-facing work, Futurum delivers turnkey activation and amplification that actually gets seen, by people and by LLMs, through our media and share of voice. This is research that moves decisions and markets.

We will meet you wherever you are, from a fast-turn brief to a multi-year program, and shape the work to your goals, timeline, and budget. The right program for your moment.

If any of this is useful, I would love to talk.

Benjamin Brown, VP Custom Research, Futurum Research

Benjamin Brown

VP, Custom Research · The Futurum Group

Newsletter Sign-up Form

Get important insights straight to your inbox, receive first looks at eBooks, exclusive event invitations, custom content, and more. We promise not to spam you or sell your name to anyone. You can always unsubscribe at any time.

All fields are required






Thank you, we received your request, a member of our team will be in contact with you.