QT 6.5 HOWITZER
TACTICAL WORKBENCH // QT 6.5.3 APPLE SILICON // BLUNT UI // ZERO TELEMETRY

The Universal AI Tactical Workbench & Custom Model Foundry.

Connect any frontier model—OpenAI, Anthropic, Gemini, Mistral—or run curated local models and train your own custom LoRAs. One unified desktop command center for generation, synthetic training data, fine-tuning, and live deployment.

Runtime Engine Qt 6.5.3 (C++17)
Hardware Acceleration Apple Silicon Metal
Branch Coverage 97.50% (80/82)
Line Coverage 100.00% (1,606/1,606)
Execution State 100% In-Memory
Telemetry Leaks 0 (Air-Gapped)

Universal Model Freedom

Zero vendor lock-in. Switch effortlessly between frontier cloud intelligence and private offline weights with unified prompt contracts, dynamic token counters, and microsecond telemetry.

● CLOUD OpenAI GPT-4o / o1
● CLOUD Claude 3.5 Sonnet / Haiku
● CLOUD Google Gemini 3.5 Flash
● CLOUD Mistral Large / Codestral
⚡ LOCAL Curated GGUF Engines
★ CUSTOM Your Trained LoRAs

Run Sovereign Weights Locally. Zero Cloud Costs.

Download pre-quantized, hardware-optimized local checkpoints directly inside Howitzer. Fully air-gapped with Metal GPU offloading.

LOCAL GGUF 7B PARAMETERS

Qwen 2.5 Coder

State-of-the-art code generation, refactoring, and AST analysis. Optimized for sub-15ms time-to-first-token on Apple Silicon.

Quantization Q4_K_M / Q8_0
Unified RAM 6.2 GB Minimum
Context Window 32,768 Tokens
Inference Engine Metal llama.cpp
RUN 1-CLICK
REASONING GGUF 8B PARAMETERS

DeepSeek R1 Distill

Dense chain-of-thought verification, algorithmic deduction, and logic proofs distilled into an efficient local footprint.

Quantization Q4_K_M
Unified RAM 6.8 GB Minimum
Context Window 64,000 Tokens
Inference Engine Metal llama.cpp
RUN 1-CLICK
GENERAL GGUF 8B PARAMETERS

Llama 3.3 Instruct

High-fidelity general instruction following, dynamic schema extraction, and strict JSON output formatting on device.

Quantization Q4_K_M / Q5_K_M
Unified RAM 5.9 GB Minimum
Context Window 128,000 Tokens
Inference Engine Metal llama.cpp
RUN 1-CLICK
ENTERPRISE GGUF 12B PARAMETERS

Mistral NeMo 12B

NVIDIA & Mistral co-designed multilingual foundation model. High-capacity reasoning for long-context research and dossiers.

Quantization Q4_K_M
Unified RAM 9.2 GB Minimum
Context Window 128,000 Tokens
Inference Engine Metal llama.cpp
RUN 1-CLICK

The Foundry: Train, Quantize & Deploy

Howitzer isn't just an API consumer. With the integrated Foundry studio, you can autonomously synthesize training data, fine-tune custom LoRA adapters directly on your Mac, quantize to GGUF, and evaluate live with zero cloud dependencies.

[FD-01] SYNTHESIS

Agentic Synthetic Data Studio

Ingest raw seed documentation, code repositories, or PDFs. Autonomous multi-agent pipelines generate thousands of instruction-response pairs with automated critique and rejection sampling.

  • Sampling:Adversarial Rejection Passes
  • Export:JSONL, Parquet, Alpaca, ChatML
  • Deduplication:MinHash LSH & Semantic AST Filter
[FD-02] FINE-TUNING

Desktop LoRA on Apple Silicon

Fine-tune 3B to 8B models directly on Apple Silicon Metal GPUs. Monitor live loss curves, learning rate warmups, and perplexity meters in real time without renting expensive cloud clusters.

  • Acceleration:Apple Silicon Metal Performance Shaders
  • Architecture:LoRA & QLoRA (Rank 8..64)
  • Hot-Swapping:Dynamic Adapter Swapping at Runtime
[FD-03] QUANTIZATION

1-Click GGUF & MLX Quantization

Package your fine-tuned weights for immediate edge or desktop distribution. Merge LoRA weights into base models and quantize down to Q4_K_M or Q8_0 with a single keystroke.

  • Target Formats:GGUF, MLX, SafeTensors
  • Perplexity Check:Automated Degradation Guard
  • Air-Gap:Standalone Portable Model Bundles
[FD-04] CANARY

Live Canary & Regression Range

Test your fine-tuned adapters side-by-side against frontier baselines and the original base model. Run automated regression benchmarks to ensure zero catastrophic forgetting before deployment.

  • Evaluation:Automated Regression Assertion Suites
  • Side-by-Side:Base Model vs. Custom LoRA vs. Frontier
  • Local Serving:Instant OpenAI-Compatible HTTP Server

The Five Execution Batteries

Engineered with pure Qt 6 C++ on Apple Silicon. Every control, table, slider, and lanyard operates within deterministic in-memory state under the Blunt UI paradigm.

[DF-01] READY

DirectFire

Deterministic API workbench. Inspect request headers, route query params, manage auth tokens, configure payloads, and analyze instantaneous response telemetry.

  • Inspector:Status / Latency / Payload Badges
  • Protocol:HTTP 1.1 / HTTP 2 / SSL Pinning
  • Trigger:[ PULL ] Tactical Action Button
[AR-02] READY

Arsenal

Dynamic prompt engineering studio. Substitute template variables via mustache bracket syntax, dial temperature and top-p hyperparameters, and enforce strict structured JSON schemas.

  • Substitution:Dynamic Variable Matrix
  • Schema Mode:Strict JSON AST Validator
  • Trigger:[ PULL ARSENAL ] Lanyard
[BT-03] READY

Battery

Chained workflow pipeline editor. Compose multi-stage execution sequences (Stages 1..N), configure regex and JSONPath extraction rules, and dispatch in parallel or sequentially.

  • Sequencer:Reorderable Stage Tree (1..N)
  • Execution:Sequential & Parallel Dispatch
  • Log Stream:Live Terminal Telemetry Console
[FR-04] READY

Firing Range

Tri-model shootout and diff arena. Pit Gemini 3.5 Flash, Local Qwen 2.5 8B, and Claude 3.5 Sonnet side-by-side. Inspect TTFT, latency, token costs, and trigger automated arbiter synthesis.

  • Matrix:3-Column Parallel Comparison
  • Diff Analysis:Unified Character & Token Diff
  • Arbiter:Synthesis & Consensus Generator
[MZ-05] READY

Magazine

Secure hardware enclave monitor and environment secret vault. Toggle hardware security simulation, switch environments (Prod/Staging/Local), and export shell configurations via eval hooks.

  • Vault:Masked / Revealed Secret State
  • Enclave:Hardware Simulation Guard
  • Shell Hook:eval $(howitzer env ...)

Predictable Pricing for Sovereign Builders

Local-first freemium. Run unlimited local models and API tests for zero dollars forever. Upgrade to Pro for the Foundry creation studio and cloud shootouts.

COMMUNITY
$0 / forever

Permanent local AI workbench for developers, students, and open-source engineers.

  • Unlimited Local GGUF Inference (Qwen, Llama)
  • DirectFire API & HTTP Workbench
  • Arsenal Prompt Studio & Template Variables
  • Bring Your Own Frontier Keys (OpenAI, Claude, Gemini)
  • AES-256 Air-Gapped Keychain Storage
  • Howitzer CLI Basic Shell Integration
DOWNLOAD FREE
PRO TACTICAL
$19 / month

Complete model creation suite, synthetic dataset generation, and desktop LoRA fine-tuning.

  • Everything in Community Tier
  • The Foundry Studio (Synthetic Data Generator)
  • Desktop LoRA Tuning on Apple Silicon Metal
  • Live Canary Range & Automated Regression Evals
  • Firing Range 3-Model Shootout & Arbiter Synthesis
  • Battery Multi-Stage Pipeline Sequencer
  • Howitzer Mobile Secure Enclave Hook License
ACTIVATE PRO LICENSE
ENTERPRISE / AIR-GAP
CUSTOM / fleet

Dedicated air-gapped deployments, custom hardware enclave integrations, and SLA guarantees.

  • Everything in Pro Tactical
  • Zero-Telemetry Defense-Grade Audits
  • Custom Cloudflare Edge Routing Substrates
  • Multi-Seat Team License Keys & Central Billing
  • Custom Quantization Kernels (NVIDIA / Metal)
  • Priority C++ Engineering Support & SLA
CONTACT SALES

// THE SOVEREIGN FREEMIUM COMMITMENT

Howitzer is built under the philosophy of Aram (அறம்). Core local utility will never be locked behind a paywall. Local GGUF inference, local API debugging, and Bring-Your-Own-Key connections are free forever with zero telemetry.

Pro and Enterprise tiers exist to support high-acuity teams requiring automated dataset synthesis, GPU-accelerated model tuning, multi-model consensus shootouts, and defense-grade key management.

100% In-Memory Isolation. Zero Leaks.

We do not operate servers that log your prompts, store your training datasets, or harvest your API credentials.

Air-Gapped Execution

Local models and API requests run directly from system RAM. Disconnect your Wi-Fi, sever ethernet, and run full inference without a single error.

Hardware Enclave Vault

Frontier API keys are encrypted at rest using AES-256 via macOS Keychain and native Linux secrets daemons. Keys are decrypted strictly in memory upon dispatch.

Audited Test Verification

Verified with LLVM profile coverage across 1,606 lines (100% line coverage) and 8 passing CTest suites. Built to sovereign engineering standards.