Industry News

Microsoft Developed AI Chip to Reduce the Cost of Running Generative Artificial Intelligence

Views : 120
Update time : 2023-05-03 14:55:41
        Dylan-Patel, principal analyst at global semiconductor research firm SemiAnalysis, recently said that OpenAI could spend up to $700,000 per day to run ChatGPT because it runs on expensive computing infrastructure.
 
 
        Whether it's writing a cover letter, generating a lesson plan, helping users optimize their profiles, or analyzing things based on facts or assumptions, ChatGPT requires a lot of computing power to provide feedback based on user input, which comes from expensive servers, Dylan-Patel said.
        Both Dylan-Patel and his colleague Afzal-Ahmad argue that while it may cost hundreds of millions of dollars to train the big language model behind ChatGPT, its operating costs or the content production behind it will be much higher, even with any reasonable deployment size that far exceeds its training costs.
        Microsoft is rumored to be developing an AI chip codenamed "Athena" to reduce the cost of running generative AI models. The report says the project has been in production since 2019 and is available for testing by a small group of Microsoft and OpenAI employees.
Microsoft previously reached a $1 billion investment agreement with OpenAI that requires OpenAI to run its models only on Microsoft's Azure cloud servers. This follows news that shortages have led Microsoft to ration GPUs for some internal teams, and that NVIDIA's processors sell for a premium, so Microsoft expects to run them more cheaply for the same workload.
        In addition to powerful performance, Nvidia's chips have significant software advantages, with most AI workloads designed for them and decades of developer experience. Microsoft currently has about 300-plus employees working on the chip.
        Sources said the chip could be released as early as next year for internal use by Microsoft and OpenAI, to which Microsoft did not officially respond, but whether it will also be used by Azure customers is still under discussion. Google has developed its own line of AI chips, TPU, and is currently the only competitor chip developing LLM, while Amazon has its own alternative product line, Trainium.

 
Related News
Read More >>
LDK220 LDO Voltage Regulators Specifications, Features, Pinout, and Applications LDK220 LDO Voltage Regulators Specifications, Features, Pinout, and Applications
Feb .02.2026
The LDK220 series of low-dropout linear regulators (LDOs) is a high-performance device designed specifically for low-power consumption and high-precision voltage regulation, widely used in scenarios such as consumer electronics, industrial control, and po
Xilinx Spartan®-7 FPGA Family: A High-Performance and Energy-Efficient Solution for Mid-Range FPGAs Xilinx Spartan®-7 FPGA Family: A High-Performance and Energy-Efficient Solution for Mid-Range FPGAs
Jan .20.2026
Xilinx Spartan®-7 FPGA Family stands as a defining solution in the mid-range FPGA landscape, blending high performance, energy efficiency, and cost-effectiveness to redefine versatility for industrial, IoT, and consumer electronics applications. Built on
Altera FLEX Series: Architecture, Innovation, and Application Across Four Generations Altera FLEX Series: Architecture, Innovation, and Application Across Four Generations
Jan .07.2026
The Altera FLEX series was more than a lineup of FPGAs—it was a blueprint for how programmable logic devices could evolve to meet diverse market needs. The FLEX 8000 laid the architectural groundwork, the FLEX 10K redefined functionality with embedded mem
LM4765 vs. LM4766: A Comprehensive Comparison of Dual-Channel Audio Power Amplifiers LM4765 vs. LM4766: A Comprehensive Comparison of Dual-Channel Audio Power Amplifiers
Dec .16.2025
Among TI standout offerings, the LM4765 and LM4766 are dual-channel amplifiers designed to cater to diverse audio needs—from compact setups to high-fidelity systems. While sharing the same product lineage, these chips differ significantly in power output,