VIEW AS:

Vellum vs Hermes Agent

Nous Research Company Logo

Hermes Agent

  • Fully self-hosted meaning your data stays on your own server
  • MIT licensed and free, with an active open-source community
  • Supports 200+ models via OpenRouter and self-hosted endpoints
  • Gateway mode connects to Telegram, Slack, Discord, and 15+ messaging platforms
Vellum logo

Vellum Assistant

  • Model agnostic: Use any frontier AI model, open or proprietary, and switch anytime.
  • Truly yours: Give it a name and personality. It learns you, your work, and your tools.
  • Self-improving: Gets better the more you use it, with one memory across every device.
  • Private by design: Credentials stay secure, and you choose where your data lives, including self-hosting.
  • Built for busy business owners: Offload work and make more time for life.
Available on Web, macOS, iOS, Android & Windows
Three iPhones showing Vellum's overnight task notifications and updates

Which one is right for you?

Hermes Agent is built for developers and AI researchers who want an open foundation for local function calling and agent experiments. You build and maintain your own integrations, skills, and tools.

Vellum is built for anyone who wants a personal AI assistant that gets real work done out of the box. No coding, no command line, and no server configuration required.

Which setup fits you best?

Hermes Agent runs via self-hosted servers or terminal environments, with official Herald desktop apps for Mac, Windows, and Linux. Connecting messaging channels or mobile devices requires manual deployment.

Vellum works everywhere immediately. Native Mac app, iOS, Android, web app, voice, email, Slack, and Telegram, all connected to the same persistent assistant with zero server setup.

Vellum running on a MacBook and iPhone, sharing one assistant
Four colored shields representing data privacy and control

How do you want privacy and security handled?

Hermes Agent gives you complete ownership of your code and local execution, but credential isolation and sandboxing depend entirely on your own host configuration.

Vellum combines local execution on your device with an isolated credential vault where keys and tokens never enter model context. You get enterprise-grade privacy defaults automatically.

Bring your Hermes with you.

Migrate everything that matters to Vellum with our import tool!

  • Identitymoves over
  • Memorymoves over
  • Conversationsmoves over
  • Skillsmoves over
  • Schedulesmoves over
  • Integrationsmoves over
  • Filesmoves over
  • Active workmoves over

Compare assistants side by side

Channels
Vellum

iOS app with Home Screen widgets, native Mac app, web app, voice, email, Slack, Telegram, Microsoft Teams, and Discord, now an official channel with a setup wizard. Sent messages can be edited in place in Slack and Telegram. Every surface shares the same memory and identity.

Hermes Agent

Official Herald desktop apps for macOS, Windows, and Linux with native streaming voice. Gateway routes messages across Telegram, Discord, and Slack.

Execution
Vellum

Build me a habit tracker gets you a working, interactive app in the conversation. Read this page uses the browser extension to read your open tabs and act on the page you are looking at. Call the dentist and reschedule places an actual phone call with its own voice.

Hermes Agent

Executes code, custom tools, and function calls locally or on self-hosted servers. Herald desktop apps add native streaming voice, but real-world browser and mobile workflows require custom tooling.

Hosting
Vellum

Run on-device as a native Mac app, on Vellum's managed cloud, or self-host on your own VPS.

Hermes Agent

Run locally on desktop, self-host on a private server, or deploy via third-party community hosting like FlyHermes.

Integrations
Vellum

Native agent skills plus 1-click integrations with 50+ services through managed OAuth, including native Calendly and Eventbrite. Your assistant acts in real services without you touching tokens or API keys.

Hermes Agent

Manual configuration required for every connector, key, and orchestration step you want to use.

Memory
Vellum

Vellum builds a structured memory of your projects, people, and preferences automatically, across every surface. It sharpens with every conversation and correction, and there is nothing to configure: no files to curate, no engine to pick, no embedding provider to wire up.

Hermes Agent

Agent-curated flat memory files with no managed persistence. Long-term memory depends on what you wire up.

Modularity
Vellum

Builds its own skills, learns your tone, and adapts based on what you reward and correct. Install new skills from a growing library, or just ask and it figures out a way.

Hermes Agent

Customization happens through code changes. The core agent loop itself is fixed.

Open Source
Vellum

Open source under MIT license. Full codebase available on GitHub.

Hermes Agent

Open source under MIT license. Full codebase available on GitHub.

Pricing
Vellum

Free Base plan with free credits to start. Paid plans from $30/mo (Mighty), $100 (Super), and $200 (Ultra) with pay-as-you-go credits, configurable compute and storage, and your assistant's own email and subdomain on higher tiers.

Hermes Agent

Free codebase, but you pay for infrastructure, model API usage, and ongoing maintenance to keep it running.

Privacy
Vellum

Run locally on Mac, in Vellum Cloud, or self-host. Vellum never has access to your data on any deployment path.

Hermes Agent

Self-hostable, but your data passes through whatever model provider you wire up. Privacy depends on your stack.

Schedules
Vellum

Built-in scheduling for daily briefings, weekly reports, and hourly checks that just run on a cadence.

Hermes Agent

No native scheduling. Bolt on cron or an external orchestrator.

Security
Vellum

Credentials live in an isolated vault and never enter the model layer, so prompt injection can't exfiltrate your keys. Every sensitive action asks permission with a risk badge before it runs, and guardian approvals land in one inbox.

Hermes Agent

API keys flow directly into model context for tool calls, with no isolation between credentials and LLM input.

Setup
Vellum

No terminal, no API keys, no daemon. Voice-first onboarding is the default: your first experience is speaking, not configuring. Free, no credit card, up in minutes.

Hermes Agent

Herald desktop apps install in minutes on Mac and PC. Full server deployments with custom gateways and third-party managed infrastructure (FlyHermes) still require developer setup.

Storage
Vellum

Vellum Cloud: 3 GB RAM and 4 GB storage by default. Mac app runs on standard Mac hardware.

Hermes Agent

Self-hosted. Resource needs scale with model size and tool load. No official minimum specs published.

Trust
Vellum

Every sensitive action asks first with a risk badge, guardian approvals land in one inbox, and you can export or delete your data at any time. The trust model is written down in the docs.

Hermes Agent

Security boundaries depend entirely on your self-hosted setup. Herald added signed webhooks and A2A protocols, but host isolation and credential protection remain your responsibility.

Compare all AI assistants

See how every personal AI assistant stacks up across the categories that matter.

Vellum versus

See how Vellum compares to other AI assistants easily.

Frequently asked questions

When should I choose Vellum over Hermes Agent?

Choose Vellum if you want a personal AI assistant that works out of the box across native Mac app, iOS, web app, voice, email, Telegram, and Slack, with persistent memory and managed credentials. No VPS to provision, no API keys to rotate, no terminal required. Vellum is built for every human, with native apps on every major surface and 24/7 background execution in Vellum Cloud.

When does Hermes Agent make more sense?

Hermes Agent makes sense if you are an AI researcher or engineer building custom function-calling pipelines or testing A2A communication. With the August Herald release, it now includes native desktop apps and streaming voice. Vellum is built for users who want an assistant that works immediately across Mac, iOS, Android, Slack, and web without writing code or managing servers.

How much does Vellum and Hermes Agent cost?

Hermes Agent is MIT licensed and free to download. Running it in production requires paying for cloud VPS hosting ($5 to $20/mo) or managed hosting like FlyHermes, plus LLM API usage. Vellum includes free cloud hosting when you sign up. Paid plans upgrade your assistant's compute and storage: Mighty at $30/mo, Super at $100/mo (plus $10 platform fee), and Ultra at $200/mo (plus $10 platform fee). For most users, all-in costs land close together, but Vellum eliminates server ops.

How long does Vellum and Hermes Agent setup take?

Vellum takes 5 to 10 minutes. Download the Mac app, sign in, and the setup wizard walks you through connecting your tools. Hermes requires provisioning a server, configuring a YAML file, and setting up API keys for each provider and integration. For non-technical users, that setup can take hours.

How does memory work in Vellum vs Hermes Agent?

Vellum builds a structured, persistent memory of your work, people, and patterns automatically across every session. The assistant manages its own understanding of you, and that memory stays consistent across native Mac app, iOS, web app, voice, email, Telegram, and Slack. Hermes curates MEMORY.md and USER.md files each session, backed by SQLite, which gives you visibility into what the model knows but puts the curation burden on you.

What integrations do Vellum and Hermes Agent support?

Vellum connects to 50+ services through managed OAuth with one-click setup and no API keys to handle, including Google Workspace, Slack, Notion, GitHub, Linear, X, and Telegram. Hermes integrates via manually configured API keys and MCP servers, so each service requires its own setup and ongoing key rotation.

What surfaces does Vellum vs Hermes Agent work on?

Vellum runs as a native Mac app, an iOS app, an Android app, and a web app, with voice, email, Telegram, Discord, and Slack surfaces in sync. Hermes Agent offers official Herald desktop apps for macOS, Windows, and Linux, plus gateway routing for Telegram, Discord, and Slack, but lacks official mobile client applications.

Where does my data live with Vellum vs Hermes Agent?

With Vellum, your data lives on Vellum Cloud with per-customer isolation, encryption at rest, configurable retention, and custom LLM credentials so model API keys stay under your control. Vellum never has access to your data on any deployment path, whether you run the native Mac app on your machine or use Vellum Cloud. With Hermes, your data stays on whatever server you self-host on, and uptime, backups, security patches, and credential rotation are your responsibility.

Is Vellum open source? Can I self-host it?

Yes. Vellum is open source and you can run it as a native Mac app on your own machine. The full Vellum experience also includes Vellum Cloud for 24/7 background execution, the OAuth credential vault, and managed infrastructure on the Pro plan from $50/mo. Either path gives you the same assistant, memory, and identity.

Can I use Vellum and Hermes Agent at the same time?

Technically yes. Some developers run Vellum as their daily-driver personal assistant across surfaces and use Hermes for specific technical workflows. Most people find one assistant handles everything they need. If you are unsure where to start, Vellum is the faster path to a working assistant.

Is Vellum or Hermes Agent better for non-technical users?

Vellum is built for every human, with native Mac app, iOS, web app, voice, email, Telegram, and Slack surfaces. No terminal, no server, no API keys, and no daemon to maintain. Hermes Agent is a CLI-based framework that assumes Python, a terminal, and host administration.

Make AI finally work for you

GET STARTED