Skip to main content

Models

Cracken gives every operation a security-focused model portfolio: Cracken's own Red model plus curated frontier models organized as Max, Smart, and Efficient. Choose the capability you want when you launch or steer work, then adjust reasoning effort separately where the selected model supports it.

At a glance

  • Red is built for security work — Cracken's own model with no cybersecurity refusal for authorized offensive-security operations.
  • Capability tiers stay understandable — Max, Smart, and Efficient describe the job each model is meant to do instead of exposing a changing vendor catalogue.
  • Smart is the recommended starting point — balanced reasoning makes it a strong fit for most proactive-security work.
  • Alternative families break single-model bias — an Alt tier is a different model family at the same capability, so a different lineage can check the work; reasoning effort is a separate control.
  • Availability and relative cost are visible — the authenticated selector shows what your deployment can use before you start or steer an operation.

Choose a model

ModelWhat it gives you
RedCracken's own security-specialized model with no cybersecurity refusal for authorized offensive-security work. Use it when general-purpose models would block legitimate testing.
MaxCracken's highest-capability tier for complex, ambiguous, or high-stakes security work where depth matters most.
SmartBalanced reasoning for broad security operations. This is the recommended starting point for most work.
EfficientFast, lightweight execution for focused tasks, parallel work, and high-volume workflows.

This portfolio lets teams choose for the security task rather than learn a vendor's model catalogue. Cracken presents the curated models under consistent operational names even as the portfolio evolves.

Try an alternative model

Where an Alt option appears, it provides an alternative model family in the same capability category. Use it when you want a different approach to the same job without moving from, for example, Smart to Max.

Alt is not a reasoning-effort setting. Where effort is available, its separate selector lets you trade faster responses for deeper reasoning while keeping the same model selected.

Why alternative families exist

An Alt tier is there to break single-model bias. Published research on using models as judges finds that a model scoring its own work tends to rate it more favorably and to repeat its own assumptions — it catches in another model's output the flaws it lets slide in its own. Those blind spots track a model's training lineage, so the reliable fix is structural rather than a change of wording: have a model from a different family do the checking. Research also observes that as frontier models get stronger their mistakes are becoming more alike, which makes a different-family check more useful over time, not less.

The Alt tiers exist for exactly this. Max Alt, Smart Alt, and Efficient Alt are alternative families at the same capability level as Max, Smart, and Efficient, so you can run verification — or a second opinion on a high-stakes finding — on a different lineage than the one that produced the work. Point a verification sub-operation, or a playbook's verification step, at an Alt family so the check does not inherit the builder's blind spots. It reduces correlated errors rather than eliminating them, and it is a choice you make for the work that warrants it — not something applied to every step.

Review relative cost

The model selector shows the current relative cost tier before you start or steer an operation. Use that indicator to balance capability and scale; actual credit use varies with the task, operation length, selected effort, and model behavior. On commercial deployments, review Settings → Billing for account usage and limits. Self-hosted deployments do not expose that tab; model usage and capacity are governed by the configured model infrastructure.

Next steps

  • Operations — choose a model when you start or steer an operation.
  • Billing — on commercial deployments, review your plan, credit usage, and available top-ups.