Skip to main content
Browse Models

TheDrummer

Behemoth 123B v1.2

Released

2024-11-30

Family

Mistral

Type

Fine-Tuned Model

Downloads

External Download

You are about to open a link to an external source. Verify the URL before continuing.

Download

4-bit GGUF (Q4_K_M)

GGUF · Behemoth-123B-v1.2-Q4_K_M.gguf

Model Report

Overview

Behemoth 123B v1.2 is a large-scale generative AI language model developed by BeaverAI, presented as a specific release within the Behemoth series. Designed for advanced conversational capabilities, rich prose generation, and immersive role-play scenarios, Behemoth 123B v1.2 emphasizes depth, creativity, and nuanced character interaction. The model has been made accessible to the wider research and open-source community, fostering collaborative evaluation and specialized application across creative and dialog-based tasks.

Futuristic space station illustration associated with Behemoth 123B v1.2

Figure 1. Illustrative branding for Behemoth 123B v1.2.

Technical Architecture

Behemoth 123B v1.2 operates with 123 billion parameters, categorizing it as a language model with a substantial parameter count for its release period, according to its Hugging Face documentation. The model leverages the BF16 tensor format, which allows efficient mixed-precision computation while maintaining numerical stability in large language model training and inference. Distributed in the Safetensors format, the model ensures secure serialization and integrity of weights, which is particularly important for research reproducibility and deployment across diverse environments.

The architecture supports conversational reasoning, creative prose generation, and sustained multi-turn dialogues. Behemoth 123B v1.2 has been optimized for tasks that demand intricate narrative understanding and improvisational capacity, with special tuning towards role-playing and character-driven applications.

Behavioral Characteristics and Use Cases

Users and early community feedback indicate that Behemoth 123B v1.2 produces varied and natural language outputs, exhibiting distinct characteristics from previous iterations in the series such as v1.1 and the 2.x series. The model demonstrates a propensity for generating less predictable, more original text, particularly in scenarios that involve narrative creativity, improvisation, and character interaction. This capacity is notable in role-playing (RP) contexts, where adherence to character profiles and dialogue realism are vital. Reports suggest that the model avoids impersonation, instead reliably adhering to the scenario or character definitions provided by users.

Behemoth 123B v1.2 is frequently employed for interactive fiction, character chat, and creative writing tasks. Users have reported that the model facilitates greater depth and realism in role-play scenarios, contributing to consistent story-driven experiences.

Anime-style character thanking user

Figure 2. Illustration reflecting community engagement and character-driven output associated with the model.

Dialogue Formatting and Chat Templates

A notable feature of Behemoth 123B v1.2 is its support for creative chat templates, most notably the Metharme (Pygmalion in ST) template, which has been described as enabling varied and imaginative conversation flows. Users have found the model effective with first-person role-play formats that interleave action, dialogue, internal thought, and narration, such as the *action* Dialogue *thoughts* Dialogue *narration* structure. This supports immersive, multi-layered exchanges essential to advanced role-playing applications.

Character cards, such as "Audrey" by thecooler from CharacterHub, are often paired with the model for defining roles and attributes. Such structured prompts enhance the consistency and depth of character interaction, aligned with community best practices for interactive story generation.

Performance Evaluation and Model Comparison

According to user observations and available community leaderboards, Behemoth 123B v1.2 exhibits observed differences in performance compared to its predecessors. Compared to v1.1, v1.2 generates outputs characterized by less predictability and greater nuance. It also exhibits conversational stability as distinct from the 2.x series models. Users have provided feedback on the writing quality, dialogue realism, and adherence to character cards or scenario constraints.

Leaderboard chart showing Behemoth model rankings

Figure 3. Leaderboard screenshot displaying rankings among Behemoth and other large language models.

REMEMBER THE CANT stylized text

Figure 4. Textual graphic associated with Behemoth 123B v1.2.

Distribution, Accessibility, and Community

Behemoth 123B v1.2 has been publicly released via the Hugging Face model repository, with variants including GGUF-optimized versions and a small-quantization version by iMatrix. The model's open distribution has encouraged active community feedback, with user engagement facilitated through the BeaverAI Discord server. As of the latest reported data, Behemoth 123B v1.2 has seen hundreds of downloads. Its download count indicates an active presence in research and enthusiast domains.

The model supports deployments in environments compatible with BF16 tensor operations and the Safetensors serialization, contributing to compatibility with contemporary machine learning frameworks. While the license associated with Behemoth 123B v1.2 is not explicitly stated in the original documentation, its public release aligns with principles of open research and reproducibility.

External Resources

For further information and ongoing developments related to Behemoth 123B v1.2, the following resources are recommended:

About Mistral: The Mistral family of AI models, developed by Paris-based Mistral AI, includes the original 2023 Mistral 7B release, as well as the more recent Mistral Small, Nemo, and Large weights.

More in the Mistral Family

Mistral AI /

Mistral Large 2

123 billion parameter model from Paris-based Mistral AI, significantly more capable than its predecessor in code generation, mathematics, reasoning, multilingual support, and function calling.
Mistral AI /

Mistral Small (2409)

A 22B parameter enterprise-grade small model, a convenient mid-point between Mistral NeMo 12B and Mistral Large 2. This version delivers significant improvements in human alignment, reasoning capabilities, and code over the previous version.
Mistral AI /

Mistral Small 3.2 (2506)

A 24-billion parameter multimodal model featuring improved instruction following, function calling, and reduced repetition over its predecessor.
Mistral AI /

Mistral Small 3.1 (2503)

A 24-billion parameter multimodal transformer supporting text and vision tasks with 128K token context length under Apache 2.0 license.
LatitudeGames /

Harbinger 24B

A 24-billion parameter language model fine-tuned on Mistral Small 3.1 Instruct, specialized for interactive storytelling and text-based adventures.
Mistral AI /

Devstral Small 1.0

A 23.6B parameter coding assistant finetuned for agentic software engineering tasks with 128K context window and 46.8% SWE-Bench performance.
Mistral AI /

Mistral Small 3 (2501)

A 24-billion parameter instruction-tuned language model with multilingual capabilities, 32K context window, and optimized low-latency inference performance.
TheDrummer /

Cydonia 24B v2

A fine-tuned 23.6 billion parameter Mistral-based model designed for long-context conversations and maintaining narrative coherence across extended dialogues.
Cognitive Computations /

Dolphin 3.0 Mistral 24B

A 24-billion parameter instruction-tuned model built on Mistral architecture with deliberately removed content filters to maximize user control over outputs.
Mistral AI /

Mistral NeMo 12B

A 12B parameter multi-lingual model that supports function calling built in collaboration with NVIDIA and trained using the new Tekken tokenizer. By some metrics, it is state-of-the-art in its size category. NeMo was trained with quantisation awareness, enabling FP8 inference without any performance loss.
TheDrummer /

Rocinante 12B v1.1

A 12.2 billion parameter text generation model optimized for creative storytelling, role-playing scenarios, and adventure-based interactive fiction applications.