Choose the assistant setup you need.

Basic supports one everyday assistant. Pro keeps work, home, or project assistants separate and unlocks GPU compute for custom voice and compatible open-weight models.

Basic

$5.99/mo + usage

A capable everyday assistant on one private cloud computer.

  • One active assistant
  • Voice, chat, research, and connected services
  • Android alpha invitations as the rollout opens
  • Review important changes before they happen
  • Set a monthly cap for managed model usage
  • Automatic compute-budget shutdown · coming soon
  • Connect your own Voice Studio over Tailscale
  • Email support

Pro

$19.99/mo

Separate assistants, add custom voice, or run compatible open-weight models on your GPU.

  • Everything in Basic
  • Separate memory and connections for each assistant
  • Keep one assistant available while another handles a project
  • Choose CPU or GPU compute for each assistant
  • Voice Studio deploys automatically on eligible GPU launches
  • Run compatible open-weight models, including Gemma variants
  • Prioritized email support

Team

$49.99/user/mo

Shared billing and controls for organizations using several assistants.

  • Everything in Pro
  • Shared workspace and billing
  • Team administration controls
  • Priority support
Setup, handled

A recommended cloud provider, location, and computer are preselected. Review or change them before launch. No command line or cloud console required.

Cancel renewal anytime. To close everything, the “Cancel and delete account” option terminates every assistant when your paid period ends, preventing further assistant compute and storage charges.

Why choose Pro?

Pro is useful when one assistant should not carry every context or connection.

Separate work from home

Give each assistant independent memory, connected services, and settings.

Keep one assistant available

Let a separate assistant handle a longer project without replacing the first.

Put a GPU behind your assistant

Deploy Voice Studio or run compatible open-weight models such as Gemma on supported GPU hardware.

What you may want to know first.

Why does every assistant get a separate cloud computer?

A separate cloud computer keeps each assistant's memory, tools, and connections isolated. Basic supports one active assistant. Pro supports several or a GPU assistant, and every running assistant creates its own compute bill.

Can I use my own model account?

Yes. Add supported model credentials and pay the provider directly, or let Pneum.ai handle model billing. Basic adds a 1% service charge to managed model usage; Pro includes managed billing without that added fee.

How does custom voice work?

On Pro, choosing a GPU VM automatically deploys Voice Studio. You can also run Voice Studio elsewhere and connect it to any assistant through Tailscale.

What else can I run on a Pro GPU?

Alongside hosted Voice Studio, Pro GPU agents can run compatible open-weight models such as Gemma. The model sizes and speeds available depend on the selected GPU, its memory, provider capacity, and region. Local inference uses the agent's GPU rather than managed-model requests, while compute and storage charges apply.

Is the Android app included?

The Android app is currently in internal release. Paid Basic and Pro customers will be invited to the alpha when downloads open, and we will email account holders when it is ready. The browser experience is available in the meantime.

Can I see the infrastructure cost before launching?

Yes. Pneum.ai preselects a recommended provider, location, and VM, then shows the estimate before deployment. You can keep the defaults or change them.

Can I cancel anytime?

Yes. Cancel renewal and keep access through the current paid period. If you want to close everything, the “Cancel and delete account” option also terminates every assistant when the paid period ends, preventing further assistant compute and storage charges.

Can I cap model spending?

Yes. When you use managed model billing, set a monthly spending cap in billing settings and review how much remains before the cap resets. Optional automatic shutdown at a compute budget is coming soon.