ANGELOSEGRETO/gpt-oss-20b

TEXT GENERATIONPricing:Input $0.3 / Output $1.2Concurrent Unit Cost:1Model Size:20BQuant:FP8Context Size:32kPublished:Sep 13, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The ANGELOSEGRETO/gpt-oss-20b is a 21 billion parameter open-weight model from OpenAI's gpt-oss series, designed for powerful reasoning and agentic tasks. It features configurable reasoning effort, full chain-of-thought access, and agentic capabilities like function calling and web browsing. Optimized for lower latency and specialized use cases, this model can run within 16GB of memory due to MXFP4 quantization.

Loading preview...

ANGELOSEGRETO/gpt-oss-20b: An OpenAI Open-Weight Model

The gpt-oss-20b is a 21 billion parameter model from OpenAI's gpt-oss series, designed for robust reasoning and agentic applications. It is released under a permissive Apache 2.0 license, allowing for broad experimentation, customization, and commercial deployment. This model is specifically optimized for lower latency and specialized use cases, capable of running within 16GB of memory thanks to MXFP4 quantization of its MoE weights.

Key Capabilities & Features

  • Configurable Reasoning Effort: Users can adjust the reasoning level (low, medium, high) to balance response speed and analytical depth.
  • Full Chain-of-Thought Access: Provides complete visibility into the model's reasoning process, aiding debugging and increasing trust in outputs.
  • Agentic Functionality: Includes native support for function calling, web browsing, and Python code execution.
  • Fine-tunability: The model can be fine-tuned for specific use cases, with gpt-oss-20b being suitable for fine-tuning on consumer hardware.
  • Harmony Response Format: Designed to be used exclusively with OpenAI's harmony response format for correct operation.

Ideal Use Cases

  • Applications requiring powerful reasoning and agentic capabilities.
  • Scenarios where lower latency and efficient local deployment are critical.
  • Custom projects benefiting from a permissive Apache 2.0 license and fine-tuning options.