🤫husshhussh
🤫husshhusshOnePuppy
🤫 Puppy One Max · Founder's Edition

Puppy One keeps your most sensitive context close. You still hold the keys.

It thinks locally whenever it can, and reaches the cloud only when you permit it. You own the device, the keys, the data, and the routing policy. This is not a laptop we resell - it is the reference engineering machine behind a private-AI system: provisioning, a curated local model pack, encrypted local RAG, hybrid-compute routing, and a transparent cost-and-energy accounting for every token it generates.

Run the Buy vs Rent calculatorThe Puppy 100 series ladder
The decision

Yes - one unit, bought now, as the reference and Founder's Edition machine.

  • ·An exceptional portable private-AI workstation - not the right default machine for every Puppy One customer.
  • ·This first unit is a Hushh-owned development, demonstration, and benchmarking machine - not the start of a resale-inventory financing strategy. Apple's retail terms are for end-user purchase, and 0% installment credit is real consumer credit risk, not a wholesale facility.
  • ·Commercial units route through an authorized Apple business procurement relationship, or the customer buys and owns the Mac directly - never bulk retail-card purchases intended for resale.
The right machine for the right workload

Where it is genuinely excellent, acceptable, or the wrong tool.

Excellent fit

  • • Private RAG across financial documents, email, CRM exports, meeting notes, research, and contracts
  • • Advisor and family-office meeting preparation
  • • Financial-statement and tax-document extraction
  • • Drafting client communications where source data should stay local
  • • Compliance pre-review and policy comparison, with human approval
  • • Private-repository coding agents
  • • Local multimodal document and screenshot understanding
  • • Speech transcription, embeddings, semantic search, and local knowledge graphs
  • • Offline and travel use
  • • Single-user and small-team interactive AI
  • • LoRA / QLoRA adaptation of small and medium models

Acceptable, with limits

  • • One heavy 70B-120B model session
  • • Two simultaneous medium-model sessions
  • • Moderate long-context analysis
  • • Local image generation
  • • Experimental fine-tuning of 30B-class models
  • • Temporary internal API serving

The wrong machine

  • • Training foundation models from scratch
  • • High-concurrency multi-tenant SaaS
  • • A 24/7 production service with a contractual uptime requirement
  • • CUDA-only research and production stacks
  • • Large-scale reinforcement learning
  • • Large generative-video batches
  • • Hundreds of parallel agents
  • • Dense models above the practical 128GB memory envelope
  • • Any workload requiring redundant nodes and automatic failover

A future Mac Studio or redundant desktop configuration is the right shape for an always-on Puppy Station. This machine is the portable, owner-grade Puppy.

BYO everything

BYO everything - compute, model, memory, infrastructure, intelligence, and above all, subscription.

Puppy One and Agent One are built to bring your own: your own compute (this machine, a cloud instance you already pay for, or both), your own model choice and the memory/context budget it needs, your own cloud infrastructure and identity stack, your own intelligence provider under your own keys - and above all, your own subscription. If you already pay for Apple One, Google One, ChatGPT Plus, Claude Pro, or a committed cloud credit, the router burns that quota down first. We are never the vendor standing between you and a service you already pay for - we are the orchestration and consent layer on top of it.

The four BYO pillars →

Quota attainment - the honest upsell signal

When a customer's existing subscription quota keeps running out - the ChatGPT Plus message cap hit every week, the Claude Pro allowance gone by Wednesday, committed cloud credits exhausted early - that consistent pattern is the signal to offer owned compute, never a sales push based on guesswork. This is how the business model meets the customer where they actually are: start on what they already pay for, and let their own measured usage tell us when a Puppy purchase is the honest next step.

The hybrid-compute doctrine

Private by default. Your own subscription next. Metered cloud last.

Private

Financial information, family data, confidential documents, private repositories, regulated workflows.

No prompt or source data leaves the machine. Local models only.

Balanced

Work that starts local and occasionally needs more headroom.

The local model handles retrieval, preparation, redaction, and ordinary reasoning. When local capability, context, concurrency, or latency is exceeded, the router burns down a subscription quota you already pay for before it ever reaches for a metered route - and only the minimum necessary context is transmitted, every route logged and visible to you.

Maximum Intelligence

Non-sensitive work where the frontier model is materially better.

An approved frontier cloud model - your own subscription or API keys first, a metered route only if nothing you already pay for has headroom - used for non-sensitive work or explicit customer-approved use, with estimated token cost and data boundary shown before execution.

The router

Seven questions, in order, every single time.

01

Data sensitivity

Is cloud transmission permitted at all?

02

Model capability

Can a certified local model meet the task-quality threshold?

03

Memory & context

Will the model and its KV cache fit safely in the certified envelope?

04

Latency & concurrency

Is the local queue within the service target?

05

BYO subscription quota

Does an existing subscription you already pay for - a message allowance, a storage tier, a committed cloud credit - have headroom to burn down before anything metered?

06

Economics

If nothing you already own or pay for has headroom, which metered route has the lower expected cost for an accepted result?

07

Consent

Has the customer permitted this specific route?

08

Audit

Record the model, route, data boundary, estimated energy, and dollar cost - every time.

Cloud becomes the default for

  • • Foundation-model training
  • • CUDA-dependent software
  • • High-concurrency workloads
  • • Large public or synthetic batches
  • • Very long contexts beyond the local certified envelope
  • • Frontier-model quality requirements
  • • Multi-tenant production services
  • • Workloads requiring high availability or geographically distributed serving
Transparency, by design

The Puppy Compute Passport

In design

Every certified model publishes a passport, not a marketing number: exact model and hash, license and permitted commercial use, quantization, runtime and version, macOS version, maximum certified context, memory consumption, prefill speed at 1K/8K/32K input tokens, decode speed at 128/512/2,048 output tokens, time-to-first-token (P50/P95), 1/2/4-session concurrency, idle/prefill/decode wall power, watt-hours per million output tokens, hardware-plus-energy cost per million tokens, four-hour thermal stability, a 72-hour service soak, and whether a given task ran local or used cloud burst. Every figure states plainly whether it is measured or estimated.

Energy and token economics

How to measure it honestly

The adapter's wattage is not inference power. A defensible number requires: battery already charged, fixed display brightness, the same macOS and runtime version, the same prompt/context/output length, separate idle/prefill/decode measurements, and at least three runs after thermal stabilization. The right units are tokens per joule, tokens per watt-hour, watt-hours per million output tokens, and million accepted tokens per kilowatt-hour - accepted tokens, not merely generated ones, because a token the user rejects is wasted compute.

The figures below are early community benchmarks on comparable M5 Max 128GB hardware - directional, not yet a Hushh-measured Compute Passport. They will be replaced by our own wall-power and runtime measurements before launch.

Model, quantizationCommunity decode rateIllustrative $ / million tokens
Llama 3.1 8B, 4-bit117 t/s~$2.10 hardware + electricity
Qwen3 MoE 30B, 4-bit176 t/s~$1.40 hardware + electricity
Gemma 4 26B-A4B, 4-bit151 t/s~$1.63 hardware + electricity
GPT-OSS 120B, 4-bit~114 t/s~$2.16 hardware + electricity
Qwen3-VL 32B dense, 4-bit27 t/s~$9.11 hardware + electricity
Local ownership vs. cloud rental

Local ownership versus cloud rental - two different comparisons

Same architecture (local Apple silicon versus a hosted Apple-silicon instance) and same business outcome (local versus a data-center GPU running the same open-weight model) are not the same comparison. Modeled at an illustrative local cost near $0.89/productive hour (capital plus electricity, three-year life, eight productive hours a day), a comparable hosted Apple-silicon cloud instance runs roughly 7x more per allocated hour, and a data-center GPU rental runs roughly 4-5x more per rented hour at public list price - cited for comparison only, never as a claim that we operate or partner with any provider named. In practice the first cloud route we reach for is never a fresh, metered account - it's the quota you already burn down under BYO Subscription.

Local wins for repeated, private, interactive, low-concurrency work. Your own existing subscription wins next, for as long as it has headroom. Metered cloud is the last resort, reserved for real parallelism a single interactive user rarely uses enough of to capture its economics.

Before this ships

Ten gates, honestly tracked.

Reference unit purchasedIn progress
Full benchmark suite run on the actual reference unitIn pursuit
Authorized commercial procurement path operational (or customer-owned-hardware route live)In pursuit
Exact wall-power measurements replace the planning assumptionIn pursuit
Five launch workflows pass quality and privacy validationIn pursuit
Curated model pack license-reviewedIn pursuit
72-hour reliability soak passedIn pursuit
Customer ownership, encryption keys, recovery credentials, and cloud consent unambiguousIn design
Warranty, returns, and replacement process documentedIn design
Every launch benchmark reproducibleIn pursuit

One is a product of Hushh Technologies Corporation (brand: 🤫 “hussh”), an independent company. One runs on third-party silicon, systems, and cloud; all company names are used solely to describe the platforms on which One software runs. Hushh Technologies is not affiliated with, endorsed by, sponsored by, or partnered with any company named.

Own the machine. Own the routing policy.

This page will be replaced, line by line, with our own measured Compute Passport as the reference unit is benchmarked - not the other way around.

See the Puppy 100 seriesThe Puppy One catalog

One is a product of Hushh Technologies Corporation (brand: 🤫 “hussh”), an independent company. One runs on third-party silicon, systems, and cloud; all company names are used solely to describe the platforms on which One software runs. Hushh Technologies is not affiliated with, endorsed by, sponsored by, or partnered with any company named.

Products

  • Agent One
  • The 🤫 One app
  • Puppy One
  • Which Puppy is right for you?
  • The Puppy 100
  • Tag One
  • The 🤫 Store
  • The 🤫 One Card
  • Pricing
  • Claim your One
  • The product roadmap

🤫 Yellow Pages

  • The 🤫 Yellow Pages
  • Discover in the feed
  • Find a local expert
  • Coverage & markets
  • Connect - in Agent One
  • Ping an expert

Business & Enterprise

  • 🤫 for Business
  • Small & medium business
  • 🤫 Concierge (VVIP)
  • 🤫 for the Enterprise
  • Industry solutions
  • Federal government & agencies
  • 🇺🇸 Defense & national security
  • For advisors (RIAs)
  • Partner Portal
  • One for Sellers
  • Developers

Watch, read & learn

  • The media library
  • The 🤫 Feed
  • See it in a minute
  • Listen - the podcasts
  • Blogs
  • Research & papers
  • Guides - by topic
  • Academy
  • The Heartbeat
  • Wiki

Company & open

  • About
  • Team
  • Investors
  • Fund A
  • Building in the open
  • Newsroom & press
  • Release notes
  • Careers
  • Contact
  • Explore - the whole site, mapped
  • Sitemap

Trust, rights & gratitude

  • The Hussh Protocol (PCHP)
  • Day 0 Trusted Circle
  • The case - a right, made enforceable
  • Data-rights landscape
  • Accessibility
  • 🤫 Champions of the Community
  • 🤫 Faculty - the professors
  • Gratitude - people we admire
  • The 1024 - humans of the world
  • Search every page
  • Browse (developer view)
🤫husshhusshKirkland, WAPrivacyTerms

© 2026 Hushh Technologies Corporation - an independent company.