analyze anything. automate everything. from your terminal.
Magine AI - Terminal-Styled AI Orchestration Platform
Available in the Chrome Web Store Magine - Spawn vision-enabled AI agents autonomously browsing the web | Product Hunt Get the Add-on for Firefox
The browser extensions are completely optional - they only help sync your signed-in sessions with Magine.
- What is Magine?
- Product Overview
- The Story of Magine
- Hayai Vision and the AGP Workflow
- How Magine Works
- Features
- CatCode Coding Agents
- Terminal Commands
- AI Browser Agents
- REST API and Webhooks
- Pricing
- Security and Privacy
- Roadmap
Magine is a retro-terminal web interface for deep GitHub profile analysis, AI-powered browser automation, and scheduled agent workflows - all controlled from a single command-line-inspired UI.
Think of it as your personal command center: type a GitHub username and get an instant deep dive on a developer, or spin up autonomous browser SDAs (Sight-Driven Agents) that can see, navigate, interact, and report back - on a schedule or in real time. Send scheduled posts to LinkedIn, get a summary of your X (Twitter) feed, triage your Gmail inbox, let a coding agent open a pull request on your repo, or automate any web task you can iMagine.
Prefer buttons to a command line? Flip the Human | Machine switch and every command becomes a form - see Human Mode.
Magine AI Architecture Pipeline SDA Browser Sessions - Agents that can see
Left: the Magine architecture pipeline Β· Right: SDA browser sessions - agents that can see
Hayai computer vision model doing its task AGP tracing web elements - the Autonomous GUI Pilot
Left: the Hayai computer vision model doing its task Β· Right: AGP tracing web elements - the Autonomous GUI Pilot
iMagine a world where AI agents can actually see.
Most AI agents today are blind as a bat π¦. They rely on APIs, structured selectors, and DOM scraping - meaning the moment a website changes a class name or moves a button, everything breaks. That fragility is why autonomous browser automation has been stuck in demo mode for years.
Magine was born to fix this. Instead of teaching agents to parse HTML, we built Sight-Driven Agents (SDAs) - autonomous browser agents that literally see the screen. They take real-time frames of the page, feed them to our vision models, plan their next move based on what they observe, and then act - just like a human would.
Cats see things humans miss. Just as cats perceive movements invisible to us, Magine's SDAs perceive web interfaces that traditional agents cannot navigate - login walls, CAPTCHAs, dynamic pages, and visual content with no API. Cats are independent, observant, and self-sufficient. So are Magine's SDAs - which is why we call them catbots.
An SDA creates Action Streams: a continuous loop of frame capture β vision-model planning β GUI and API actions executed. Every step is recorded, so you can scrub through an SDA's work frame by frame.
- They see, not scrape - SDAs work from what is on screen, so they survive redesigns, canvas apps, and pages with no API
- They learn - agents keep short-term memory within a run and long-term memory across runs, and improve from their own successes and failures
- They are isolated - every user runs in a private, sandboxed browser environment with its own cookies, storage and fingerprint
- They are fast - browsers boot lazily, live views stream without lag, and long runs compact their own context to keep token costs down
- Mixture of Experts - Magine's native models, cloud models and specialised vision models run in parallel for faster, more accurate decisions
- GitHub Productivity Tracker - beyond browsing, Magine analyses developer profiles with AI-driven scoring and market-value estimation
SDAs already power real workloads on providers like Zeupiter - from automated cost management to Gmail triage to full vibe deployments: describe what you want in plain English, and an SDA plans, tests, and ships it.
Every SDA can drive a page one of three ways, switchable per agent:
| Workflow | How it sees the page | Best for |
|---|---|---|
classic |
Reads the page structure | Simple, well-behaved sites |
cdp |
Talks to the browser directly - faster, lighter | Speed on standard sites |
agp |
Looks at the screen - no page code at all | Canvas apps, heavily scripted sites, anything a DOM-based agent can't use |
AGP (Autonomous GUI Pilot) is Magine's DOM-free workflow. At every step it captures a clean frame of the page and hands it to Hayai Vision, Magine's own computer-vision model, trained to find every interactive control on a screen in a single pass - buttons, fields, links, toggles, menus - the way a person scanning the page would. AGP numbers those controls on the frame and asks the planner one narrow question: which one, and what action? The answer lands on real screen coordinates, so nothing depends on class names, selectors, or a site's markup staying still.
- Works where DOM agents can't - canvas apps, PDF viewers, custom widgets, and pages that actively fight automation
- Degrades gracefully - if the detector is unsure, AGP falls back to classical vision + OCR so the step still completes
- Gets better with every run - successful and failed actions feed back into Hayai Vision's training, so the model improves from real pages rather than static datasets
- Transparent - the run log tells you which detector handled each step
catbot create sda <prompt> # a new vision-first agent catbot workflow <name> agp # switch an existing agent to the vision workflow
π How Hayai Vision was built - why interface perception is not ordinary object detection, the data flywheel behind the model, and where it is heading: read the story at zeupiter.com/research .
βββββββββββββββββββββββββββββββββββββββββββββββββββββββ
β magine terminal β
β βββββββββββββββββββββ β
β > analyze torvalds β
β β
β Fetching GitHub data... β
β Running AI analysis... β
β βββββββββββββββββββββββββββββββββββββββββββββ β
β β β
Profile Score: 98/100 β β
β β π Top Languages: C, Shell, Perl β β
β β π₯ Contribution Streak: 4,021 days β β
β β π‘ AI Insight: "The most prolific..." β β
β βββββββββββββββββββββββββββββββββββββββββββββ β
β β
β > catbot create "daily-monitor" β
β Agent created. Use `catbot task` to assign work. β
β β
β visitor@magine:~$ _ β
βββββββββββββββββββββββββββββββββββββββββββββββββββββββ
- Type a GitHub username, a task, or a command - or click it in Human mode
- Watch SDAs work in the live viewer, frame by frame, or take the controls yourself
- Review results in
catbot-output.log, ask follow-up questions, and schedule it to run again
iMagine knowing any developer in seconds
- Instant analysis - type any GitHub username to get a comprehensive profile breakdown
- AI-powered insights - deep analysis of contributions, repos, and coding patterns
- Profile scoring - 0β100 rating based on activity, impact, and community engagement
- Beautiful cards - generate shareable, embeddable GitHub profile cards with 30 themes - no rΓ©sumΓ© building
- Your timezone, not the server's - Most Active Hour and Busiest Time are shown in your own local time
iMagine an army of cats browsing for you
Each agent gets its own real cloud browser inside a sandbox that is yours alone. Agents can:
- Browse autonomously - navigate real websites using the
classic,cdporagpworkflow - Show their work live - watch every step in the SDA Live Viewer
- Hand you the controls - open the agent's live screen, take over with your own mouse and keyboard, then release
- Take natural-language tasks - "go to LinkedIn and check my notifications"
- Work on any site - Gmail, LinkedIn, YouTube, X, Amazon, and anything else with a screen
- Ask for credentials safely - prompt-based auth when a login is needed, never stored on disk
- Replay every run - frame-by-frame review with thumbnail navigation
- Use your files - reference uploads with
#report.pdfand they are attached to the run
Quick tasks - just tell CatBot what to do:
catbot do "check NVIDIA stock price on Yahoo Finance" catbot do "order Sony WH-1000XM5 headphones from Amazon and pay using my card" catbot do "send an email on Gmail to Arthur about yesterday's meeting" catbot do "search arXiv for latest transformer-based LLM papers from 2026" catbot do "what's the latest Veritasium video about on youtube" catbot do "apply to senior frontend developer jobs on Indeed" catbot do "research the best Mumbai street food on Reddit" catbot do "check what's happening on my X (Twitter) feed"
For work you want to keep, name and save an agent with catbot create.
iMagine your agents running while you nap
- Natural language scheduling - say "every weekday at 9am" instead of writing cron expressions
- AI cron parser - LLM-powered conversion of human language to precise schedules
- Preset schedules - hourly, daily, twice daily, weekdays, every 6h/12h
- Heartbeat mode - reactive watchers that fire the moment a page changes (see below)
- Timezone-aware - schedules respect your local timezone
iMagine an agent that wakes up only when something happens
Heartbeat agents are reactive, not scheduled. Instead of running on a timer, the agent keeps an eye on a page and reacts within seconds of a real change - a new email, a status flip, a fresh notification. Only then does the full agent wake up with your action prompt and spend regular agent tokens.
# Minimal - `every` and `url` are OPTIONAL catbot create heartbeat \ watch "new unread email" \ do "summarise it in 2 lines and reply 'on it' if it asks a question" # With url (recommended) and a custom safety-net interval catbot create heartbeat \ watch "new unread email at the top of the inbox" \ do "open the newest unread email, summarise it in 2 lines, and reply 'on it' if it asks a question" \ every 60s \ url https://mail.google.com/
Units:
every <n>s,every <n>m,every <n>h, orevery [5 minutes]. A bare number is seconds. Range: 10sβ5m.
- Cost: 2 π± (vs 1 π± for a regular agent) - a browser tab stays alive continuously
everyis a ceiling, not a timer - default 60s; it only forces a check for changes the watcher can't see (canvas, video, cross-origin frames). Real changes fire fasterurlis optional but recommended - the watcher returns to that page if the tab drifts- Quote
watchanddo- bare text is rejected so a typo can't silently become the action - Coalescing - three changes during one running action produce one wake-up that handles all three
- Backoff - transient errors back off (30s β 1m β 5m β 15m β 60m) so a flaky page can't drain your wallet
Use it for inboxes, dashboards, notification feeds, queue UIs, status pages, ticket boards - anything where the right moment to act is "whenever something new shows up".
iMagine Magine without the command line
A Human | Machine switch in the top-right of the terminal turns every command into a labelled button or a small form: creating an SDA becomes a name and a description; scheduling becomes a dropdown. Nothing is left out - the same tabs cover SDAs, CatCode, your account and, for administrators, the whole platform. It works before you sign in too (visitors get a Welcome tab with sign-in, registration and the free developer card), and every button runs the exact command the terminal would, so the two modes can never disagree. Flip back to Machine any time for the full command line, autocomplete and history.
iMagine your perfect terminal aesthetic
- Light and dark modes - clean light theme inspired by GitHub, plus 30 card themes
- Voice commands - click the paw button or type
voiceto speak commands - Draggable panels - arrange your workspace with resizable terminal windows
- Command history - arrow keys to navigate previous commands;
#autocompletes your files,@your agents - Tab status light - the browser tab shows a yellow dot while an SDA is working and a green one when it finishes while you're away, so you know when to come back
- Mobile responsive - full experience on phones and tablets
iMagine your CatBot is already logged in
Available in the Chrome Web Store Get the Add-on for Firefox
The Magine Bridge extension (Chrome, Edge, Brave and Firefox) shares the sites you are already signed in to with your agents, so they skip the login wall. It is entirely optional.
- Per-domain consent - you pick each site to sync, or share everything you're signed in to with one click
- Single-use pairing codes that expire in 60 seconds
- Encrypted at rest (AES-256-GCM) with a 7-day auto-expiry
- Keep in sync - optional automatic refresh every 6 hours
- Cascade revoke -
browser unlinkdeletes every session for that profile
iMagine unlimited analysis power
- Free tier - every visitor gets tokens to start analysing
- Token consumption - different actions cost different amounts
- Top-up - purchase additional tokens when you need more
- Blurred previews - see what premium analysis looks like before buying
iMagine secure access everywhere
- Local accounts - register with username/password
- Google OAuth - one-click sign in with Google
- QR login - scan a QR code from your phone to log in on desktop
- Session management - secure token-based sessions
iMagine an agent that clones your repo, proposes a plan, and opens the pull request
CatCode gives you isolated project workspaces and a coding agent that works on real repositories - from the terminal, in the in-browser editor, or bridged to your own machine.
catcode clone https://github.com/you/repo my-project # private repos use your own GitHub token catcode plan Add rate limiting to the API # the agent drafts a plan - no files touched yet catcode approve # the agent builds the approved plan on a branch catcode diff my-project # review every change it made catcode pr Add rate limiting # push the branch and open the pull request
- Plan first -
catcode planshows the summary, steps and files the agent intends to touch; nothing changes until youapprove. Usecatcode doto plan and build in one step - Never on main - the agent works on its own branch and ships through a pull request you review
- Two identities, on purpose - cloning uses your GitHub token (from
credentials), so private repos stay private to you; pushing and opening PRs uses Magine's bot identity. Add the bot as a collaborator on any repo Magine should open PRs against - Rich editor -
catcode editopens a full in-browser code editor with live autosave; files in your Cattery open the same way and export to PDF or PPTX in one click - Build and run -
catcode buildandcatcode run <cmd>stream output live intocatbot-output.log - Drag and drop - drop files into the Cattery or a CatCode workspace to upload them; download any workspace as a zip
- Smart routing - mention "cattery", "workspace" or "project" in an agent prompt and generated files land in the right place
Browsers can't touch your local files, run your tests, or compile your code. catcode-cli bridges that gap: planning and code generation stay in the Magine cloud, while terminal-bound commands (tests, builds, git) run locally on your machine through a loopback-only bridge that never accepts remote connections and asks you in your own terminal before running anything destructive.
# Install npm install -g catcode # or, straight from Magine curl -fsSL https://magine.cloud/api/catcode/install | bash # Start the bridge in your project directory catcode start catcode start --port 4411 --dir . # custom port or directory
| Command | Description |
|---|---|
<username> |
Type any GitHub username to generate an embeddable SVG profile card |
analyze <user> |
Deep-dive analysis - score, heatmap, stack breakdown |
analyze --refresh |
Refetch (purge profile cache) |
help |
Show full command list with descriptions |
about |
The story behind Magine |
docs |
Open the documentation page |
clear |
Clear terminal output |
exit |
Close the terminal (mobile) |
| Command | Description |
|---|---|
theme |
Change card theme (30 presets) |
customize |
Customize card colors (bg, text, accent) |
social |
Add social media links to your card |
bio |
Set your job title and bio (supports [text](url) markdown links) |
preview |
Preview current card settings |
preview <opt> |
Toggle card sections: ai, activity, deep, social, devworth |
| Command | Description |
|---|---|
login |
Login with username or email |
login google |
Sign in with Google |
register |
Create a new account |
forgot |
Reset password (via GitHub token) |
logout |
Sign out of your session |
whoami |
Show current user info |
credentials |
Update your GitHub token / AI key |
tokens |
Check your token balance and usage |
topup |
Purchase token packages (Dodo Payments / PayPal) |
request <msg> |
Request credits, report bugs, or share feedback πΎ |
| Command | Description |
|---|---|
catbot create <prompt> |
Create a new AI browser agent (1 π±) |
catbot create sda <prompt> |
Create a vision-enabled SDA agent (1 π±) |
catbot create heartbeat watch "<w>" do "<a>" [every <n>[s|m|h]] [url <u>] |
Create a reactive heartbeat agent (2 π±) |
catbot list |
List all your agents with status |
catbot task <id|name> <task> |
Assign a natural language task to an agent |
catbot run <id|name> |
Run a CatBot (agent or SDA) |
catbot do <prompt> |
Quick one-off browser task (always starts fresh) |
catbot continue |
Resume a previously paused quick task |
catbot do stop |
Cancel a running quick task |
catbot stop <id|name> |
Force-stop any agent - cancels run, watcher loop, cron, all modes |
catbot schedule <id|name> <schedule> |
Set a recurring schedule (natural language, preset, or cron) |
catbot heartbeat <id|name> watch "<w>" do "<a>" [every <n>] [url <u>] |
Convert an existing bot to a reactive heartbeat |
catbot heartbeat <id|name> off |
Disable heartbeat (revert to agent mode) |
catbot mode <id|name> <agent|sda|heartbeat> |
Switch CatBot mode |
catbot workflow <id|name> <classic|cdp|agp> |
Switch browsing workflow - agp is the Hayai Vision workflow |
catbot prompt <id|name> <task> |
Set or change the agent's task prompt |
catbot rename <id|name> <new> |
Rename a CatBot (also re-routes @<name> tags) |
catbot isolate <id|name> on|off |
Wall a bot off from your shared cookie jar |
catbot color <id|name> #<hex> |
Set the agent's identity colour in the logs |
catbot email <id|name> <email> |
Email you the results of each run |
catbot delete <id|name> |
Permanently remove an agent |
catbot memory |
View saved browsing memories |
catbot memory delete <site> |
Delete memory for a specific site |
catbot memory clear |
Clear ALL agent memories |
catbot logs / catbot stats |
Activity logs and statistics |
catbot mood / catbot treat / catbot scold |
Check the agent's state, reward it, or correct it |
| Command | Description |
|---|---|
catcode init <project> |
Create a new project (git init) |
catcode clone <url> [project] |
Clone a repo - private repos use your GitHub token |
catcode projects |
List all your projects |
catcode plan <task> |
Draft a plan for your approval - no files touched |
catcode approve |
Approve the plan and let the agent build it |
catcode do <task> |
Plan and build in one step (catcode do <proj> | <task>) |
catcode diff [project] |
Review the agent's uncommitted changes |
catcode pr [title] |
Push the branch and open a pull request |
catcode branch <name> |
Create and switch to a branch |
catcode commit <message> |
Commit all changes |
catcode push <repo-url> |
Push to GitHub |
catcode status [project] |
Show git status |
catcode files [project] |
List project files |
catcode edit [project] |
Open the rich in-browser code editor |
catcode build [project] |
Run the project build |
catcode run <cmd> [project] |
Run a shell command in the project |
| Command | Description |
|---|---|
browser |
Help and overview of the bridge |
browser install |
Show the extension download links and setup steps |
browser link |
Get a 60-second pairing code for the extension popup |
browser status |
List your linked browser profiles and synced domains |
browser unlink <profileId> |
Revoke a profile (cascade-deletes all its imported sessions) |
| Command | Description |
|---|---|
apikey create [name] |
Generate a new API key (max 5 active) |
apikey list |
List your active API keys |
apikey revoke <id> |
Revoke a key permanently |
webhook register <endpoint> <url> |
Register a webhook callback URL |
webhook test <endpoint> |
Send a test delivery |
webhook status <endpoint> |
Check webhook config |
webhook remove <endpoint> |
Remove callback URL |
| Command | Description |
|---|---|
light / dark |
Switch between light and dark themes |
voice |
Toggle voice commands (speech-to-text) πΎ |
timezone |
Show or set timezone (IANA format) |
timezone auto |
Detect the timezone from your browser |
history |
Show command history (β/β to navigate) |
history clear |
Clear saved command history |
CatBot agents are autonomous browser instances powered by SDAs. Each agent gets its own real cloud browser and can:
- Navigate - go to any URL
- Click - interact with buttons, links, and menus
- Type - fill forms, compose messages, and search
- Scroll - explore long pages
- Read - extract text and understand page content
- Screenshot - capture what it sees at every step
- Wait - pause for pages to load or auth flows to complete
- Log in - request credentials securely when authentication is needed
Create Agent β Assign Task β Agent Opens Browser β Executes Steps
β β β β
catbot create catbot task Live View (SDA) frames saved
β β β β
Schedule it NL instructions Watch in real-time Review frames later
# Using presets catbot schedule <id|name> daily catbot schedule <id|name> every_hour catbot schedule <id|name> weekdays_9am # Using natural language (AI-parsed) catbot schedule <id|name> every monday at 8am catbot schedule <id|name> twice a week on tuesday and friday catbot schedule <id|name> every 30 minutes catbot schedule <id|name> first day of every month # Using raw cron catbot schedule <id|name> 0 */6 * * *
There are no site-specific commands to learn. Anything you want done on LinkedIn, Gmail, X or any other site is just a prompt - the agent navigates the real UI with the same anti-bot stack it uses everywhere, and a single login persists across runs:
catbot do log into LinkedIn and check my notifications
catbot do search LinkedIn for "senior backend engineer" jobs in Berlin and save the top 10
catbot do post on LinkedIn: "Shipping today - neurons graph, catnips, and faster agent screens."
catbot do connect with the first 5 people on LinkedIn who match "founding engineer"
Magine provides a REST API for triggering agents and retrieving results, plus real webhooks that push results to your server.
> apikey create my-integration > apikey list > apikey revoke <key-id>
curl -X POST https://magine.cloud/api/catbots \ -H "Authorization: Bearer mk_..." \ -H "Content-Type: application/json" \ -d '{"botId": "<id>", "additionalPrompt": "optional extra tasks"}'
The additionalPrompt field appends extra tasks to the agent's base prompt (never overrides it).
curl https://magine.cloud/api/catbots?botId=<id>&limit=5 \ -H "Authorization: Bearer mk_..."
Returns paginated run results with steps, summary, tokensUsed, and structured output.
Register a callback URL and Magine will POST results to your server whenever an agent run completes:
> webhook register <endpoint> https://your-server.com/webhook > webhook test <endpoint> # Send a test delivery > webhook status <endpoint> # Check webhook config > webhook remove <endpoint> # Remove callback URL
Deliveries carry an X-Magine-Signature header (HMAC-SHA256) for verification. n8n: use a Webhook node (push) or an HTTP Request node to POST to /api/catbots (pull).
Agents can spawn other agents - mention @another-agent in a prompt and it is invoked at run time, or call the spawn webhook directly. Inline arguments use a CLI-style syntax and override body fields:
@data-extractor pull orders --since=2025εΉ΄01ζ01ζ₯ --format=csv --label="last quarter"
Every call is recorded in the π§ Neurons view - a graph of who calls whom, coloured by how reliably each agent serves its callers, with a live status dot on every agent that is running right now.
Reference any file you have uploaded by name with # and Magine attaches it to the run:
Summarise the contents of #report.pdf and compare it to #last-week.csv
Uploads stay available to the same agent's future runs (and appear in # autocomplete) until you delete them or they expire after 30 days. They are private to your account; deleting a bot wipes its files.
| Package | Tokens | Price |
|---|---|---|
| Free | 1,000,000 starter tokens | 0γγ« |
| 1 Cat | 5,000,000 tokens | 5γγ« |
| 5 Cats | 20,000,000 tokens | 15γγ« |
| 10 Cats | 50,000,000 tokens | 25γγ« |
| 50 Cats | 200,000,000 tokens | 100γγ« |
Token costs vary by action. Profile analysis uses fewer tokens than deep AI analysis or browser agent tasks. Use the topup command in the terminal to purchase more, or request to ask for free credits. π±
- API keys are stored hashed (SHA-256) and shown once; stored credentials and uploads are encrypted at rest with AES-256-GCM
- Every user is isolated - agents, browsers and CatCode workspaces run in per-user sandboxes; no account can reach another's files, even by guessing a name
- Agent credentials are held in memory only during execution and never written to disk
- Session tokens are cryptographically random with TTL-based expiration
- Passwords never reach the logs - anything typed into a password field is masked in step records
- Auto-cleanup - screenshots, agent logs, and uploaded files are garbage-collected on rolling retention windows; deleting a bot or your account cascade-deletes everything tied to it
- We do not sell or share your data - ever
- Open-source, self-hosted Magine - testing of the self-hosted build is in development and it opens to the public in a few months. Star the repo and raise an issue to follow along or volunteer as a tester.
- Claude Code plugin - drive Magine's SDAs and CatCode workspaces straight from Claude Code, shipping alongside the self-hosted release.
- Enterprise marketplaces - Magine is coming to the NVIDIA, AWS and Google Cloud marketplaces for enterprise deployment.
Built with π & obsessive attention to terminal aesthetics