Spotlight

Case Study Microsoft

How Microsoft scaled global content delivery

Find out how Microsoft used Gcore to strengthen delivery across regions.

case study ProSieben GNTM app TOPSHOT

How ProSieben scaled GNTM's app TOPSHOT

Explore how ProSieben brought real-time AI portraits to GNTM's audience.

case study Higgsfield

How Higgsfield scaled AI video generation

See how Gcore helped Higgsfield scale with GPUs and Managed Kubernetes.

case study Fawkes Games

How Fawkes Games stopped DDoS attacks

See how Gcore protected gaming servers from massive DDoS threats without disrupting gameplay.

We're hiring

Help build the next chapter of the web

We're not just filling seats. We're building a team that will write the next chapter of the internet.

  1. Home
  2. Blog
  3. New AI inference models on Application Catalog: translation, agents, and flagship reasoning
AI

New AI inference models on Application Catalog: translation, agents, and flagship reasoning

  • December 7, 2025
  • 2 min read
New AI inference models on Application Catalog: translation, agents, and flagship reasoning

We’ve expanded our AI inference Application Catalog with three new state-of-the-art models, covering massively multilingual translation, efficient agentic workflows, and high-end reasoning. All models are live today via Everywhere Inference and Everywhere AI, and are ready to deploy in just 3 clicks with zero infrastructure management.

We update our Application Catalog whenever promising new AI models are launched so that Gcore AI customers can deploy them with ease in just 3 clicks. It’s part of our commitment to making AI simple and easy to use, without compromising on performance.

Let’s see what these new models offer.

Meta/Facebook SeamlessM4T v2 Large

Meta's Facebook’s SeamlessM4T v2 Large is a massively multilingual, multimodal translation model designed to break down communication barriers across nearly 100 languages. This model is a unified system, meaning it can handle complex, real-world translation needs without relying on separate models for speech recognition (ASR) and text-to-speech (TTS).

Key capabilities:

  • Supports all major translation directions: speech-to-speech, speech-to-text, text-to-speech, and text-to-text.
  • Global: Seamlessly translates speech and text in approximately 100 languages.
  • Ideal for: Real-time voice translation services, powering multilingual contact centers, and creating sophisticated global AI assistants that can interact naturally across different linguistic modes.

Deploy Meta SeamlessM4T v2 Large

MiniMax M2

The MiniMax M2 is a powerful addition for developers focusing on agentic workflows and coding tasks. Utilizing a Mixture-of-Experts (MoE) architecture, this LLM offers 230 billion total parameters while activating only about 10 billion at any given time. This selective activation makes it exceptionally fast and cost-efficient.

Key capabilities:

  • Delivers near-flagship performance quality without the high latency and cost typically associated with models of this size.
  • Highly optimized for complex agentic use cases, enabling it to excel at planning, tool-use, and sophisticated multi-step workflows.
  • Building efficient tool-using agents, acting as a high-speed dev copilot, and automating multi-step workflow pipelines.

Deploy MiniMax M2

Qwen3-235B-A22B-Instruct-2507

Qwen’s latest flagship model is built for the most sophisticated enterprise and research applications. This large MoE model (235 billion total parameters with 22 billion active) is instruction-tuned and multilingual, providing world-class performance across challenging benchmarks. Its standout feature is its massive context window, accommodating extremely long inputs.

Key capabilities:

  • High-end reasoning: Excels in complex areas like reasoning, advanced mathematics, and sophisticated code generation.
  • Massive context: Supports context up to 262K tokens, allowing the model to manage and analyze huge documents, codebases, or extended conversation histories.
  • Ideal for: Creating sophisticated enterprise copilots for research and analysis, powering highly complex agents that require deep context memory, and advanced general-purpose text generation.

Deploy Qwen3-235B-A22B-Instruct-2507

Deploy the latest models in 3 clicks and 10 seconds

With Gcore AI inference solutions, you can eliminate the operational complexity of AI without compromising performance or power. Experience low-latency routing, transparent and efficient costs, and 3-click deployment. Simply open the Gcore Customer Portal, choose your model from the Application Catalog, and launch your endpoint.

Deploy these new AI models today

Try Gcore AI

Gcore all-in-one platform: cloud, AI, CDN, security, and other infrastructure services.

Related articles

Sifted 100 France & Benelux 2026 announcement with falling confetti and spotlights.
Gcore named in Sifted Top 100 France & Benelux 2026

Gcore has been recognized as one of the top 100 fastest-growing technology startups in France and Benelux by Sifted — one of Europe's leading tech publications. Our inclusion in the B2B SaaS & Cloud Infrastructure category points to ris

GCORE and NVIDIA's Global Inference Routing, accelerated by NVIDIA Dynamo, features a glowing green network sphere.
Gcore introduces Global Inference Routing accelerated by NVIDIA Dynamo

Earlier this year we brought NVIDIA Dynamo to Gcore — one-click disaggregated inference that delivered up to 6× higher GPU throughput and 2× lower latency inside a deployment, by separating prefill and decode and routing each request to the

Two founders discuss Melious AI moving its CDN and DNS to Gcore.
Why Melious AI moved its CDN and DNS to Gcore: a founder conversation about sovereign AI in Europe

For many startups, infrastructure decisions are mostly about performance, pricing, and developer experience. For Melious AI, they are also about trust.Melious AI is a German startup building a European AI platform around privacy, transparen

GCORE and Graphiant logos connected by an 'X', signifying a partnership.
Gcore and Graphiant: Accelerating sovereign AI infrastructure with secure neo-cloud connectivity

As enterprises move AI from experimentation into production, they face a new infrastructure challenge. AI applications, models, and data are no longer confined to a single cloud or data center. Instead, they are distributed across multiple

5 insights on AI infrastructure from Nexus Luxembourg 2026

Nexus Luxembourg is Europe's premier AI and technology summit, and this year's edition brought together more than 10,000 visitors, 150+ speakers, and 250 startups from over 50 countries. Gcore CEO Andre Reitenbach joined LuxProvide's Arnaud

An isometric illustration of a secure server rack with a shield icon and glowing data activity.
AI sovereignty isn’t politics: it’s a sales requirement

Across Europe, I keep seeing the same pattern in public sector deals, regulated industries, and anything that smells like critical infrastructure: "AI sovereignty" has moved from a nice-to-have to the first real checkpoint in the deal. Not

Subscribe to our newsletter

Get the latest industry trends, exclusive insights, and Gcore updates delivered straight to your inbox.