logo
#

Latest news with #DmitryPimenov

Amazon announces first-ever availability of OpenAI models for its cloud customers, company says, ‘The addition of...'
Amazon announces first-ever availability of OpenAI models for its cloud customers, company says, ‘The addition of...'

Time of India

time4 days ago

  • Business
  • Time of India

Amazon announces first-ever availability of OpenAI models for its cloud customers, company says, ‘The addition of...'

Amazon Web Services (AWS) has announced that OpenAI's open weight models will be available for the first time on Amazon Bedrock and Amazon SageMaker AI. This will allow the company's customers to build generative artificial intelligence (AI) apps, as this move will make OpenAI's technology accessible to millions of AWS customers. In a blog post, Amazon has confirmed that OpenAI's two new open-weight foundation models, gpt-oss-120b and gpt-oss-20b, will be offered on its platforms. According to AWS, when running in Amazon Bedrock, the larger model is three times more price-efficient than the comparable Gemini model, five times more than DeepSeek-R1, and twice as efficient as the OpenAI o4 model. This announcement reinforces AWS's commitment to providing a wide selection of models. It expands the existing range of fully managed models in Amazon Bedrock and makes them accessible in Amazon SageMaker JumpStart. What Amazon and OpenAI said about this collaboration Commenting on the latest partnership, Atul Deo, director of product at AWS, said: 'Open weight models are an important area of innovation in the future development of generative AI technology, which is why we have invested in making AWS the best place to run them, including those launching today from OpenAI. The addition of OpenAI as our newest open weight model provider marks a natural progression in our commitment to bringing cutting-edge AI to organisations worldwide, and the unmatched size of our customer base marks a transformative shift in access to OpenAI's advanced technology.' Meanwhile, Dmitry Pimenov, product lead, OpenAI, noted: 'Our open weight models help developers—from solo builders to large enterprise teams—unlock new possibilities across industries and use cases. Together with AWS, we're providing powerful, flexible tools that make it easier than ever for customers to build, innovate, and scale.' by Taboola by Taboola Sponsored Links Sponsored Links Promoted Links Promoted Links You May Like Use an AI Writing Tool That Actually Understands Your Voice Grammarly Install Now Undo Why Amazon wants to use OpenAI's open weight models Amazon is integrating OpenAI's open weight models due to their strong reasoning abilities, which are well-suited for building AI agents—an area that's reshaping how businesses operate. Through Amazon Bedrock AgentCore, companies can run these models natively at scale and with the security needed for production-level applications. They can also gain access to features like gpt-oss-120b and gpt-oss-20b, which can be integrated with Bedrock's enterprise-grade safeguards, including Guardrails that filter out up to 88% of unsafe content. Upcoming features like Custom Model Import, Knowledge Bases, and additional customisation will further enhance this integration. Meanwhile, Amazon SageMaker AI enables customers to use these open weight models with tools for pre-training, evaluation, fine-tuning, and deployment, offering a full-stack solution for developing advanced AI systems. OpenAI's open weight models are well-suited for agentic AI use cases due to their strong reasoning capabilities, have an efficient size-to-performance ratio and support instruction-following and tool use. With a 128K context window, they can handle long documents and conversations, making them ideal for tasks like coding, scientific research, and customer service. These models also include safety measures, having undergone thorough training and evaluation to support responsible AI deployment. iOS 26 Public Beta Is Here: Apple's Biggest Redesign Since iOS 7

Cerebras Helps Power OpenAI's Open Model at World-Record Inference Speeds: gpt-oss-120B Delivers Frontier Reasoning for All
Cerebras Helps Power OpenAI's Open Model at World-Record Inference Speeds: gpt-oss-120B Delivers Frontier Reasoning for All

Business Wire

time4 days ago

  • Business
  • Business Wire

Cerebras Helps Power OpenAI's Open Model at World-Record Inference Speeds: gpt-oss-120B Delivers Frontier Reasoning for All

SUNNYVALE, Calif. & SAN FRANCISCO--(BUSINESS WIRE)--Cerebras Systems today announced inference support for gpt-oss-120B, OpenAI's first open-weight reasoning model, now running at record-breaking inference speeds on the Cerebras AI Inference Cloud. Purpose-built for complex challenges in math, science, and code, this 120B-parameter model achieves intelligence on par with top proprietary models like Gemini 2.5 Flash and Claude Opus 4—while delivering unmatched speed, cost efficiency, and openness. "Through deployment partners like Cerebras, we're together able to provide powerful, flexible tools that make it easier than ever to build, innovate, and scale," said Dmitry Pimenov, product lead at OpenAI. Share For the first time, an OpenAI model leverages Cerebras' wafer-scale AI infrastructure to run full-model inference. By eliminating GPU memory bandwidth bottlenecks and communication overhead, Cerebras wafer-scale AI inference delivered a world-record 3,000 tokens per second output speed — a major advance in responsiveness for high-intelligence AI. 'OpenAI's open-weight reasoning model release is a defining moment for the AI community,' said Andrew Feldman, CEO and co-founder of Cerebras. 'With gpt-oss-120B, we're not just breaking speed records—we're redefining what's possible. OpenAI on Cerebras delivers frontier intelligence with blistering performance, lower cost, full openness, and plug-and-play ease of use. It's the ultimate AI platform: smart, fast, affordable, easy to use, and fully open.' At over 3,000 tokens/second, organizations will be able to use Cerebras-powered gpt-oss-120B to build live coding assistants, instant large document Q&A and summarization, and fast agentic research chains. These high-intelligence AI reasoning use cases have long wait times on proprietary models running on GPUs – that lag is now dramatically reduced with gpt-oss-120B on Cerebras. Developers can swap their existing OpenAI endpoints for Cerebras in 15 seconds. No refactoring. No migration headaches. Just instant access to the highest performance and quality gpt-oss-120B models running on the Cerebras Cloud. The open-weight Apache 2.0 license from OpenAI gives users full control to fine-tune for their domain, deploy on-prem for sensitive or regulated data, or move freely across clouds. "Our open models let developers—from solo builders to large enterprise teams—run and customize AI on their own infrastructure, unlocking new possibilities across industries and use cases," said Dmitry Pimenov, product lead at OpenAI. "Through deployment partners like Cerebras, we're together able to provide powerful, flexible tools that make it easier than ever to build, innovate, and scale." Experience the fastest AI inference today: Developers and enterprises can now access gpt-oss-120B on the Cerebras Cloud with a free API key ( About Cerebras Systems Cerebras Systems is a team of pioneering computer architects, computer scientists, deep learning researchers, and engineers of all types. We have come together to accelerate generative AI by building from the ground up a new class of AI supercomputer. Our flagship product, the CS-3 system, is powered by the world's largest and fastest commercially available AI processor, our Wafer-Scale Engine-3. CS-3s are quickly and easily clustered together to make the largest AI supercomputers in the world, and make placing models on the supercomputers dead simple by avoiding the complexity of distributed computing. Cerebras Inference delivers breakthrough inference speeds, empowering customers to create cutting-edge AI applications. Leading corporations, research institutions, and governments use Cerebras solutions for the development of pathbreaking proprietary models, and to train open-source models with millions of downloads. Cerebras solutions are available through the Cerebras Cloud and on-premises. For further information, visit or follow us on LinkedIn, X and/or Threads.

DOWNLOAD THE APP

Get Started Now: Download the App

Ready to dive into a world of global content with local flavor? Download Daily8 app today from your preferred app store and start exploring.
app-storeplay-store