Token Factory office hours: Build and ship with open models

Join the Nebius team for a live, developer-focused office hours session on Nebius Token Factory — our production inference platform for deploying, serving, and fine-tuning open-source and custom AI models at scale.

This is an interactive Q&A session designed to help you move faster with your AI applications. Bring your code, API requests, or architecture questions, and we’ll work through them together.

What you’ll get:

  • Live help integrating with the OpenAI-compatible API
  • Guidance on selecting the right model from 60+ available models (Nemotron, GLM, DeepSeek, and more)
  • Best practices for fine-tuning and post-training workflows
  • Performance, latency, and cost optimization tips
  • Troubleshooting support for real implementation challenges

Whether you’re building your first prototype or running production workloads, this is your opportunity to get direct, hands-on guidance from the engineers behind Token Factory.

Register to get the recording

To make the session as useful as possible, come prepared with your questions.

Helpful topics include:

  • A code snippet or API request that isn’t working as expected
  • An architecture or deployment challenge
  • Questions about model selection for your use case
  • Fine-tuning or evaluation workflows
  • Performance, scaling, or inference optimization

Try Nebius AI Cloud console today

Get immediate access to NVIDIA® GPUs, along with CPU resources, storage and additional services through our user-friendly self-service console.