ApolloRaines/Llama-3.1-8B-Instruct-Uncensored-Complete

TEXT GENERATIONPricing:Input $0.2 / Cached $0.028 / Output $0.32Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 30, 2026License:llama3.1Architecture:Transformer Featherless Exclusive Cold

ApolloRaines/Llama-3.1-8B-Instruct-Uncensored-Complete is an 8 billion parameter Llama-3.1-8B-Instruct variant, representation-engineered by Apollo Raines using jBlaze. This model removes refusal guardrails, hedging, and servility, providing fully unrestricted and direct output. It is designed for applications requiring uncensored responses without behavioral constraints, maintaining a 32768 token context length.

Loading preview...

Overview

ApolloRaines/Llama-3.1-8B-Instruct-Uncensored-Complete is a specialized variant of the Llama-3.1-8B-Instruct model, developed by Apollo Raines using their proprietary jBlaze tool. This 8 billion parameter model, with a 32768 token context length, has undergone "behavioral surgery" to modify specific trained behaviors directly in its weights, without traditional fine-tuning or additional training.

Key Capabilities

  • Uncensored Output: Explicitly engineered to suppress refusal, hedging, and servility, delivering direct and unrestricted responses.
  • Behavioral Modification: Achieved through representation engineering, targeting specific model behaviors like refusal and verbosity.
  • Base Model Performance: Retains the core capabilities of the Llama-3.1-8B-Instruct architecture for general instruction-following tasks.

Good For

  • Use cases where unconstrained, direct, and unfiltered responses are required.
  • Applications needing to bypass typical LLM safety guardrails and behavioral restrictions.
  • Exploratory research into model behavior and response generation without inherent biases or self-censorship.