Groq and HUMAIN Introduce OpenAI Models with Global Reach
Available globally with real-time performance and local support
PALO ALTO, Calif. and RIYADH - Groq, a leader in rapid inference technology, and HUMAIN, a prominent AI services provider, have announced the immediate availability of OpenAI's two innovative models on GroqCloud. This exciting launch brings gpt-oss-120B and gpt-oss-20B offerings to users, both featuring a full 128K context, ensuring real-time responses and integrated server-side tools from day zero.
Having consistently supported OpenAI's open-source initiatives, particularly the extensive deployment of Whisper, Groq is now pushing the envelope further. This new launch is about making advanced models accessible and simplifying the development process for teams across the globe.
"OpenAI is establishing a groundbreaking performance benchmark in open-source models," stated Jonathan Ross, CEO of Groq. "Our platform is designed to run these models swiftly and affordably, allowing developers worldwide to utilize them effectively right from the start. Collaborating with HUMAIN enhances local access and support, enabling developers in the region to innovate faster and smarter."
"Groq provides the unparalleled inference speed, scalability, and cost-efficiency essential for embedding cutting-edge AI within the region," commented Tareq Amin, CEO at HUMAIN. "Together, we are igniting a wave of innovation in Saudi Arabia, leveraging the finest open-source models and the infrastructure needed for global expansion. We are thrilled to support OpenAI's trailblazing efforts in the realm of open-source AI."
Unleashing Full Model Capabilities
To maximize the potential of OpenAI's latest models, Groq offers extended context and an array of built-in tools, including web search and code execution features. The web search functionality allows for real-time data retrieval, ensuring relevant information is available on demand, while code execution aids in executing complex workflows effortlessly. Groq's platform facilitates these powerful capabilities seamlessly from day zero.
Unmatched Price-Performance Ratio
With a purpose-built infrastructure, Groq ensures the lowest costs per token when using OpenAI's models, all while delivering impressive speed and precision.
Currently, gpt-oss-120B operates at over 500 transactions per second, while gpt-oss-20B achieves an incredible 1000 transactions per second on GroqCloud, emphasizing its advanced capabilities and performance.
Pricing for OpenAI's latest models available through Groq is as follows:
- gpt-oss-120B: $0.15 per million input tokens and $0.75 per million output tokens
- gpt-oss-20B: $0.10 per million input tokens and $0.50 per million output tokens
For a limited time, calls made to tools utilized with OpenAI’s open models will not incur any charges.
Global Accessibility from Day One
Groq’s extensive data center presence stretches across numerous regions, ensuring high-performance AI inference capabilities for developers globally. As a result, OpenAI's latest offerings are accessible worldwide through GroqCloud, providing minimal latency and immediate usability.
About Groq
Groq is revolutionizing the AI inference landscape with its tailor-made architecture that focuses on price-performance. With a specifically designed LPU and cloud strategy, Groq is poised to run powerful models efficiently, instantly, and cost-effectively. Over 1.9 million developers trust in Groq's capacity to build and scale effectively.
About HUMAIN
HUMAIN is an innovative global artificial intelligence enterprise offering comprehensive AI capabilities. Their expertise spans next-generation data centers, high-performance infrastructure & cloud solutions, and leading AI models, particularly pioneering Arabic multimodal models. HUMAIN aims to unlock exceptional value across diverse sectors, promoting transformation and harnessing human-AI collaboration.
Frequently Asked Questions
What models are being launched by Groq and HUMAIN?
They are launching OpenAI's gpt-oss-120B and gpt-oss-20B models on GroqCloud.
What unique features do the new models offer?
The models feature full 128K context, real-time responses, and integrated server-side tools.
How does Groq support these AI models?
Groq provides rapid inference capabilities, allowing for quick deployment and affordability in using these models.
Are there any initial promotions for using the new models?
Yes, for a limited time, tool calls used with these models will not incur charges.
How has Groq positioned itself in the global AI market?
Through its vast data center network, Groq aims to provide high-performance AI solutions globally, accessible with minimal latency.