Skip to content
  • Models
  • Rankings
  • Ori
Sign Up
Sign Up
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for InferenceNet

inference.net

Browse models provided by inference.net (Terms of Service)

3 models

Tokens processed on OpenRouter

  • Favicon for inference-net
    Inference.net: Schematron V2 TurboSchematron V2 Turbo

    Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather than through system or user prompts.

    by inference-netSep 12, 2026128K context$0.03/M input tokens$0.15/M output tokens
  • Favicon for inference-net
    Inference.net: Schematron V2 SmallSchematron V2 Small

    Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema in response_format rather than through system or user prompts.

    by inference-netSep 12, 2026128K context$0.05/M input tokens$0.23/M output tokens
  • Favicon for moonshotai
    MoonshotAI: Kimi K3Kimi K3

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at navigating large repositories, using tools, debugging, and iterating against images, logs, tests, and runtime feedback. Its architecture uses KDA and Attention Residuals for computational efficiency.

    by moonshotaiJul 16, 20261.05M context$2.10/M input tokens$10.95/M output tokens