Serenities AI™
ConnectorsMarketplacePricing
DocsArticles
Sign InGet Started

Product

  • Pricing
  • Demo & Showcase
  • Connectors
  • Builder Challenges
  • AI Models

Resources

  • Guide
  • Flows Guide
  • Quick Start
  • Seller Guide
  • Affiliate Program

Company

  • About
  • Editorial Policy
  • Privacy Policy
  • Terms of Service

Connect

  • Email Support

© 2026 Serenities AI™. All rights reserved.

  • Rankings
  • All Models
  • Providers
  • Compare
  • Models Pricing
  • Benchmarks
Providers/Inference

Inference

Inference serves 4 of the models we track. We do not support bring-your-own-key for this provider yet.

text — 4 models

ModelMade byInput / 1MOutput / 1MContext
Llama 3.2 1b InstructNVIDIA$0.01$0.0116K
Llama 3.2 3B InstructNVIDIA$0.02$0.0216K
Llama 3.1 8B InstructNVIDIA$0.025$0.02516K
Llama 3.2 11b Vision InstructNVIDIA$0.055$0.05516K

Inference — Frequently Asked Questions

How many AI models does Inference serve?

Inference serves 4 of the models we track, across 1 model type.

What is the cheapest model on Inference?

Llama 3.2 1b Instruct at $0.01 per million input tokens.

Can I use my own Inference API key?

Not yet — Inference is not currently one of the providers Serenities supports for bring-your-own-key.