Skip to content

Tool Library

Every tool in Alfrada OS, clustered exactly as they cluster on alfrada.comSpecifications, with what each tool actually does and when to reach for it. Every tool below is available to every user on day one — rates per call are published separately. Grounded in the live tool catalog — not marketing copy.

481 tools in the registry across 8 clusters. That number includes every sub-action inside MCP-backed integrations (Slack 43, GitHub 42, Facebook 34, GA4 36, Calendar 28, Twitter/X 27, Gmail 22, Canva 14, Hunter 8, Discord 10). What you see below is the 87 top-level tool cards (MCP integrations clubbed), grouped the same way the website groups them.

How To Read This Page

For each tool you get:

  • What it does — the specific capabilities, operations, filters, or presets the tool exposes.
  • When to reach for it — concrete scenarios.
  • Try asking — one or two prompts you can paste.

You don't need to name tools in your prompts. Describe the job and Alfrada OS picks the right ones. This page exists so you can calibrate what's possible before you ask.

Tool availability depends on three things

Every tool below ships to every plan on day one. A tool is usable in your session only if all three are true:

  1. The tool is enabled in your session (Tool Selector).
  2. The required account is connected (for MCP integrations).
  3. Your Safety Center settings allow the action.

Naming convention — "My {Platform}"

Clusters marked My {Platform} (for example, My Gmail, My Twitter/X, My GDrive) use your connected account. In your UI, Alfrada OS renders them as {FirstName}'s {Platform} — so John sees "John's Gmail", not "My Gmail". The canonical, backend name stays My {Platform} so logs, agent reasoning, and approval prompts are consistent across users.


1. Core & web research

Search, browse, academic, patents, images, news, search trends, job listings, rank tracking & live library docs — 14 tools.

What it does — ranked web search with snippets, titles, and links. The workhorse for any "find me information about…" request.

When to reach for it — quick lookups, disambiguation, finding the right URL before handing off to a scraper (e.g. finding a LinkedIn profile URL before LinkedIn Profile Lookup can read it).

Try asking"what's the current state of the EU AI Act?"

Google Images

Release note: Google Images — Localised Image Search That Saves Straight Into Your Session.

What it does — Google Images search via SearchAPI with country/language auto-localised from your detected location. By default the top 5 results download into the session out/ folder (set download_count to a different number, or 0 for metadata-only) and are indexed so Presentation Create, Video Editor, and Report Generator can use them in the same turn. Each saved file ships with alt text, title, source name, source link, and a ready-to-render credit line.

Filters: image size (16 options including large, medium, icon, fixed-resolution and megapixel thresholds), colour (15 — color, black_and_white, transparent, plus 12 named hues), image type (photo, clipart, line_drawing, gif, face), aspect ratio (square, tall, wide, panoramic), recency (last hour / day / week / month / year), usage rights (Creative Commons only, or commercial-use), SafeSearch (active / blur / off), and pagination.

Per-image download cap: 12 MB, 15 second timeout. Hotlink-blocked or oversized images are skipped — the rest still save.

When to reach for it — building decks or articles that need real photography rather than AI renders; news-fresh visuals; license-clean imagery (CC) for redistribution; transparent icons or line drawings for diagrams. Different from Pexels Stock (curated royalty-free) — Google Images covers the open web with optional license filtering.

Try asking"find five large landscape photos of the Brandenburg Gate at night and save them with credits" or "pull 10 Creative Commons photos of the Tokyo skyline at sunset, wide aspect".

Google News

Release note: Google News — Localised News Search That Cites Itself.

What it does — Google News search via SearchAPI with country/language auto-localised from your detected location — Berlin gets the German edition, London the British one, Karachi the Pakistani one, and so on. Top 5 articles auto-save as .news.json citations in the session out/ folder (set save_count to a different number, or 0 to skip saves) and are indexed for Report Generator and Presentation Create. Each citation carries title, source, date, link, snippet, locale context, and a ready-to-render credit line.

Article thumbnails — SearchAPI returns them as base64 data URIs — are decoded once on the backend, written to out/ as real image files, and substituted with short relative paths in the response. Thumbnails render in the chat card without inflating LLM token cost.

Filters: recency (last_hour, last_day, last_week, last_month, last_year), sort order (most_recent for chronological, default is relevance), pagination, optional location override. Per-row inline citation shows the publisher hostname (clickable) plus, for saved articles, the citation filename — so the link between "what I read" and "what got saved" is visible without scrolling to a footer.

When to reach for it — current events, breaking news, regional press coverage, news-fresh research that needs source attribution. Pairs naturally with the Fast Browser (drill into one article's full text) and Report Generator (cite saved articles in a generated brief). Different from Google Search (general web results) — Google News is specifically curated news-publisher coverage with publication dates surfaced.

Try asking"latest news about climate policy", "top headlines in Germany today", or "recent coverage of SpaceX Starship launches from the last day".

Release note: Google Trends — Ask What the World Is Actually Searching For.

What it does — Google Trends via SearchAPI, in four shapes. Interest over time returns the 0–100 relative interest index week by week (hour by hour on short ranges) plus a per-term summary Alfrada OS computes on top: average, latest, peak value with its date, and a rising / flat / falling read. Interest by region ranks countries, or drops to states/provinces, US metros (DMA) or cities when scoped to one country. Related queries and related topics each return a top list and a rising list, where values are percentage growth and Breakout means over 5000%.

Up to 5 terms compare on one shared scale in a single comma-separated query. Nine time ranges (now 1-H through all, default today 12-m) plus custom YYYY-MM-DD YYYY-MM-DD windows. Five surfaces: web (default), images, news, shopping, YouTube. Time-series and region results save as a CSV in the session out/ folder, indexed for Code Executor, Excel Ops, and Report Generator; save_data: false skips it. The chat card draws the curve with a hover crosshair and per-date tooltip, with each line labelled and every number repeated as text underneath.

When to reach for it — demand trends, seasonality, hype checks, brand or product comparisons, which markets want something most, and the long-tail terms people search alongside yours. Different from Google News (what got published) and Google Search (what's on the web) — Trends is the only tool here that measures demand. Note that it does not auto-localise to your country the way the other Google tools do: worldwide is the default, because a demand question is usually a global one. Name a place to scope it.

Market sizing works by anchoring — the agent pairs a term whose absolute volume you already know with your target in one comparison call, converts the ratio with a conversion rate and a live price, and can calibrate the whole chain against a public comparable's reported revenue. See the Demand Curve Before The Deck playbook.

Try asking"is ozempic still trending?", "compare search interest in chatgpt, claude and gemini over the past 5 years", or "which German states search most for ski holidays".

Google Jobs

Release note: Google Jobs — Every Board's Openings, With the Pay Attached.

What it does — Google Jobs via SearchAPI: live openings Google pooled from LinkedIn, Indeed, company career pages and job boards, each credited to its source and listing the other boards carrying the same role. Every listing carries what you decide on — salary range where the employer published one, full-time / part-time / contract, how long ago it was posted, remote and "no degree mentioned" flags, and the benefit chips Google attached — plus the Qualifications bullets Google extracted from the posting, and a direct apply link.

Filters are written into the query rather than passed as fields: "remote", "part time", "internship", "no degree", "since yesterday", "in the last 3 days". Up to 30 listings per search (10 by default), paged and de-duplicated for you. Every search saves jobs_<query>_<timestamp>.csv into the session out/ folder with the complete description text for each role, indexed for Code Executor, Excel Ops and Report Generator; save_data: false skips it, full_descriptions: true puts the long text in the chat too. Name a city and the search is anchored there — when Google doesn't recognise the place, Alfrada OS retries with it written into the query and says so if the results ended up country-wide.

When to reach for it — job hunting, what a role currently pays, who is hiring for a skill or in a city, remote and part-time openings, and reading a competitor's roadmap off its open roles. Also an alt-data source: for a private target or a market with no published research, live postings are the only continuous operating disclosure there is — function and location mix, net-new roles week over week, published pay, and the stack named in the requirements. Treat it as a fixed-query sample, never as a headcount count. Different from LinkedIn Jobs, which reads LinkedIn alone with applicant counts and LinkedIn's own filters — reach for Google Jobs first for coverage across every board, and drop to LinkedIn when the question is specifically about LinkedIn.

Try asking"find me remote product designer jobs", "who is hiring data scientists in New York and what do they pay?", or "part time barista jobs in Chicago posted this week".

Rank Tracking

Release note: Rank Tracking — Where a Site Sits on Google, By Keyword.

What it does — live organic Google position for a named site on up to 10 keywords, via SearchAPI. Scans the top 100 (depth 10–100) for a country or city and desktop / mobile / tablet. Pass a domain (acme.com) and you get that site's exact rank plus any competitors you name; omit it and the card lists who owns page one. Subdomains count as the same site; add a path to track one section (acme.com/blog).

Every run saves a CSV of every scanned organic row into out/ for before/after tracking (save_data: false skips it). The chat card shows best / average / how many keywords sit in the top 10, with title and URL at that rank. Ads, maps packs, and other SERP features are not counted.

When to reach for it — SEO rank checks, competitor visibility, before/after of content work, finding keywords with demand but no presence. Different from Google Search (read the web) and Google Trends (relative demand) — Rank Tracking is the position measurement.

Try asking"where does salesforce.com rank for best crm software?", "track nike.com and adidas.com for running shoes in London, on mobile", or "who owns page one for crm software in the US?"

What it does — AI-powered deep search that returns structured results with source citations and extracted answers.

When to reach for it — fact-checking, multi-faceted research, comprehensive topic analysis. When a plain Google Search isn't enough.

Try asking"give me a structured briefing with sources on the current state of AI regulation in the US vs EU vs China."

What it does — academic papers and citation data via Google Scholar — titles, authors, abstracts, citation counts.

When to reach for it — literature reviews, finding peer-reviewed sources, checking the academic grounding of a claim, citation analysis.

Try asking"find the most-cited papers from the last 3 years on transformer scaling laws."

What it does — patent filings, claims, and prior art via Google Patents. Searchable by keyword, inventor, or assignee.

When to reach for it — IP landscape analysis, prior art research, competitive technology intelligence, invention scouting.

Try asking"map patent activity from the last 3 years around RAG architectures. Who's filing and what claims?"

Browser (Fast)

What it does — reads any publicly accessible web page and downloads files from the internet.

When to reach for it — reading an article end-to-end, pulling a CSV/PDF from a URL, letting the agent inspect a specific page you linked to.

Try asking"read [URL] and summarise the key arguments."

Agentic Browser

What it does — autonomous browser that navigates pages, clicks, scrolls, fills forms, handles logins, and downloads files from protected sites.

When to reach for it — paywalled content, JS-heavy single-page apps, intranet pages, workflows that need multi-step navigation. Fallback when Fast Browser returns incomplete or blocked content.

Try asking"log into my newsletter dashboard and download the last 90 days of subscriber analytics as a CSV."

YouTube

Release notes: Video Lab Works On Your Footage — YouTube Rips, Aleph 2 Edits, Magnific 4K, Act-Two.

What it does — search YouTube by keyword, extract full text transcripts, and download footage into the session. Three actions: search (returns metadata like title, views, duration, channel), transcript (full subtitle text), and download (saves the video — 480p/720p/1080p ceiling, default 720p — or just the audio track into the session's in/ folder, ingest-ready for the Video Editor and Video Lab remixing; sources up to 20 minutes by default).

When to reach for it — finding tutorials, news coverage, product reviews, conference talks. Summarising video content or extracting quotes without watching (transcript is faster and cheaper than download for that). Ripping a clip you want to actually edit, remix, or upscale.

Try asking"find the most-watched talks on AI alignment from the last 12 months, grab transcripts, and summarise the common themes." — or "download this YouTube talk at 720p, ingest 12:00–14:30, and cut a vertical highlight clip."

Context7 — Resolve Library ID (MCP)

What it does — maps a package or framework name (e.g. "next.js", "langgraph") to a Context7 library ID. Called automatically before Query Docs when the agent only has a friendly name.

When to reach for it — you don't call this directly. It runs as a prep step inside library lookups.

Context7 — Query Docs (MCP)

What it does — fetches up-to-date documentation for the resolved library or framework. API references, current usage patterns, function signatures — live, not from training cutoff.

When to reach for it — any coding task where you want the latest docs. Library debugging, migration guides, CLI tool usage.

Try asking"how do I configure OAuth in Auth.js v5? Use Context7 for current docs."


2. Core utilities

Code, servers, GitHub, speech, plans, history, background jobs — 8 tools.

Code Executor

Release note: Daytona Can Now Operate Your Servers Without Seeing The Keys.

What it does — one Daytona execution surface with two isolated targets:

  • Python analysis — a 24-hour persistent sandbox with plotly, pandas, numpy, scipy, scikit-learn, xgboost, PDF/DOCX readers, and session files (in/, out/).
  • Server operations — runs a command on your own server using keys stored in your Vault; reads run immediately, changes ask first; it can stay connected up to 24 hours and check back on long jobs.

When to reach for it — CSV/Excel processing, charts, statistical modelling, document transforms, or data collection/devops on your own AWS, Google Cloud, Azure, Exoscale, VPS, or private server.

Try asking"read in/sales.csv, compute month-over-month growth by region, and produce an interactive plotly chart" or "on prod, start the container build, keep Daytona for 45 minutes, and wake in 10 minutes to check it."

Speech

What it does — converts text into natural-sounding audio. Two voices: Alfred (male) and Ada (female). Six acting styles — excited, thoughtful, confident, analytical, gentle, dramatic — that can be blended with inline audio tags to get pauses, emphasis, and pacing where you want them.

When to reach for it — voice-over for videos and presentations, spoken summaries, podcast intros, accessibility audio versions of reports. The output MP3 is reusable by the Video Editor for narrated-video workflows.

Try asking"narrate this paragraph in Ada's voice, thoughtful and analytical, with a pause after the second sentence."

Todo List

What it does — tracks multi-step tasks and project progress. Creates lists with titles, step titles and descriptions, and updates status as steps complete.

When to reach for it — long agentic runs (for your visibility), project tracking, smoke-testing a workflow end-to-end.

Try asking"create a to-do list for launching a newsletter: content calendar, platform setup, subscriber acquisition, measurement."

History Searcher

Release note: History Search Now Finds Files, Tool Results, And Other Conversations.

What it does — mixed search across your past conversations, uploads, generated files, and past tool results. Default search returns all four together so images cannot crowd out the chat that explains them. find_files is a fuzzy filename lookup across every session; list_files lists one session; full_conversation pages a transcript; search_in_file reads inside one large artifact. If this session is empty, the card shows the top 5 hits from other conversations, labelled as coming from elsewhere.

When to reach for it — resuming work from weeks ago, finding a specific draft or file, rebuilding context for a follow-up session. Use it before you re-explain context the agent might already have. Pass real words for a semantic search; * is for list/find/transcript paging.

Try asking"search my history for the investor deck draft I was working on last month."

Background Jobs

Release note: Long Renders Keep Going — The Chat Does Not Freeze.

What it does — inspect or stop long tools that detached from the turn in this conversation. Video Lab, Music Lab, Image Lab batches of 2+ images, heavy Video Editor renders, and long Code Executor runs return a live card instead of freezing the composer. action: "list" shows what is still running; action: "cancel" plus a job_id stops one. Reload shows running / delivering / finished / failed / cancelled rather than a forever spinner.

When to reach for it — you changed your mind mid-render, the prompt was wrong, or an output is no longer needed. Cancelling discards the result and does not refund provider spend already incurred. Not Task Scheduler (a new run later) and not Wake Me (a timer that re-enters the chat).

Try asking"what's still rendering in this conversation?" or "cancel that music job — wrong mood."

Other Workspaces

What it does — pulls a focused summary from one of your other Spaces — the themed workspaces your chats are grouped into. A chat normally sees only its own Space's context; this tool reaches across to another Space when you explicitly name it, or scans all of them, and returns extracts from that workspace's rolling summary.

When to reach for it — carrying a decision, file name, or conclusion from one workspace into the current chat without switching: "what did I figure out about X in Learning AI?"

Try asking"check my Marketing space — what tagline did we settle on last month?"

Artifact CRUD

What it does — direct file housekeeping for the current session. It reads the exact on-disk contents of anything you uploaded or Alfrada OS generated, and writes, corrects, renames, or deletes generated text files (markdown, CSV, JSON, HTML, SVG and similar). It can also copy a file from an earlier session into this one when you continue old work.

When to reach for it — mostly agent-initiated. You use it when you want a byte-accurate read of an uploaded file, a small fix inside a generated file without regenerating it, or a clean-up of the session's outputs.

Try asking"open the CSV I uploaded and show me the exact header row" or "rename report_final_v2.md to board-report.md and delete the older drafts."

My GitHub (42 tools via MCP)

What it does — reads and manages your GitHub account from chat. It can browse and search repositories and code, and read files, READMEs, branches, and tags. It works issues and pull requests end to end — list, create, comment, review, and merge, confirming with you before anything irreversible. It also tracks commits, releases, and Actions workflow runs (including triggering workflows and re-running failed jobs), and can commit single-file or atomic multi-file changes to a branch.

Full operation list
  • Auth — gh_get_user to confirm the account
  • Repos — gh_list_repos, gh_get_repo, gh_get_readme, gh_get_content (read files or list directories), gh_search_repos, gh_search_code, gh_list_branches, gh_list_tags
  • Issues — gh_list_issues, gh_get_issue, gh_create_issue (confirm first), gh_update_issue (edit / close), gh_list_issue_comments, gh_create_issue_comment, gh_search_issues (cross-repo, supports is:issue, is:pr, label:, author:)
  • PRs — gh_list_prs, gh_get_pr, gh_create_pr (head branch must exist), gh_update_pr, gh_merge_pr (confirm), gh_list_pr_files, gh_list_pr_commits, gh_create_review, gh_list_pr_reviews, gh_list_pr_comments
  • Commits — gh_list_commits (optional path/author/date filters), gh_get_commit, gh_compare_commits
  • Releases — gh_list_releases, gh_get_latest_release, gh_create_release (confirm)
  • Workflows — gh_list_workflows, gh_list_workflow_runs, gh_get_workflow_run, gh_list_workflow_jobs, gh_trigger_workflow (requires workflow_dispatch), gh_rerun_workflow, gh_rerun_failed_jobs
  • File ops — gh_create_or_update_file (single file), gh_commit_files (atomic multi-file; use base_branch when creating a new branch)

Try asking"list all open PRs in alfrada/api tagged 'bug', check their CI status, and summarise."


3. Platform & social

Social surfaces, messaging, design & people lookup — 17 tools.

What it does — searches posts, threads, and profiles by keyword or handle, with engagement stats.

When to reach for it — public sentiment, tracking trending topics, brand mentions, gathering opinions on events.

Try asking"what's the sentiment on Twitter about the new EU AI Act enforcement timeline? Sample a range of technical and non-technical voices."

Reddit Reader

What it does — reads Reddit posts, comments, and subreddit discussions by topic.

When to reach for it — community sentiment, honest product reviews, troubleshooting advice, niche topic deep-dives, grassroots opinions.

Try asking"what are r/MachineLearning users saying about the latest Claude release? Pull the top threads from the last week."

What it does — searches TikTok videos, hashtags, and creator profiles with engagement data.

When to reach for it — viral trend analysis, Gen-Z / millennial sentiment, influencer discovery, short-form video content research.

Try asking"find the top 10 trending TikTok videos about productivity apps this month, including view counts and creator handles."

What it does — searches Instagram posts, reels, and profiles by keyword or hashtag, with engagement data.

When to reach for it — visual brand analysis, influencer research, competitor social presence, aesthetic / design trend tracking.

Try asking"map how Series B DTC skincare brands are using Reels: top 10 posts by engagement, what formats work."

LinkedIn Profile Lookup

What it does — reads public LinkedIn profile data — experience, education, skills, headline. Runs on verified URLs only, so Alfrada OS first runs a Google Search to find the right profile URL. For posting on LinkedIn (not reading), use the LinkedIn — publishing integration below.

When to reach for it — professional background research, executive profiling, team analysis, recruiting intelligence.

Try asking"look up the VP of Engineering at Anthropic and summarise their career trajectory."

LinkedIn Jobs

What it does — live LinkedIn job listings, filterable by:

  • Title (e.g. "Staff ML Engineer")
  • Location (city, country, region)
  • Work type — remote, hybrid, on-site
  • Contract type — full-time, part-time, contract, internship, temporary
  • Experience level — internship, entry, associate, mid-senior, director, executive
  • Posted recency — last 24h, last week, last month
  • Keywords — free-text filters inside the job description

Returns structured rows: title, company, location, posted date, applicant count, salary range (when listed), work type, contract type, seniority, direct apply link.

When to reach for it — job market research, competitive hiring analysis, salary benchmarking, talent market intelligence.

Try asking"which Series B AI startups in Europe are hiring Staff ML Engineers this week? Rank by applicant count."

WhatsApp History

What it does — searches your personal WhatsApp message history by contact or keyword, lists recent conversations, and downloads images and files shared in chats. If a contact name isn't found, Alfrada OS lists your recent chats first to pin down the right one.

Argus guide — if you are setting up the bidirectional WhatsApp agent experience (pairing, ACL modes, group modes, wakeword behavior, trust boundaries, and rollout sequencing), read Argus On WhatsApp.

When to reach for it — finding a message someone sent you last month, retrieving a shared photo, building context from a client conversation.

Try asking"find the specs my contractor sent me in WhatsApp last Tuesday, download the PDF."

WhatsApp Send

What it does — sends WhatsApp messages from your agent's paired WhatsApp number (not your personal phone). Works with contacts or groups by name — no phone numbers needed — and attaches files, images, and voice notes: audio is automatically converted so it arrives as a real WhatsApp voice note, not a file. For reminders, it can ask the recipient to reply "done" and reliably match that reply back to the exact reminder it sent.

Argus guide — for the full bidirectional WhatsApp agent experience (pairing, ACL modes, group modes, wakeword behavior, trust boundaries), read Argus On WhatsApp.

When to reach for it — sending a generated report to a client chat, posting a summary to a group, voice-note reminders, follow-ups where you want confirmation the recipient actually did the thing.

Try asking"send the PDF we just made to the Family group on WhatsApp with a one-line summary" or "WhatsApp Sarah a reminder to take her medication at 8pm and ask her to reply done."

WhatsApp Save Contact

What it does — teaches your agent who someone on WhatsApp is. Save a name and a short note for any number or group ("Tom — my plumber", "lawyer at Acme"); every future message from them then arrives already labelled with your name and context, overriding whatever display name the sender chose for themselves.

When to reach for it — after an unknown number messages your agent, or to give standing context on important contacts so replies are informed.

Try asking"save that number as 'Tom the plumber' and note he's quoting on the bathroom."

My Slack (43 tools via MCP)

What it does — reads, writes, and manages a connected Slack workspace. It can catch up on channel history and threads, search messages workspace-wide, send richly formatted messages and DMs, edit or delete its own posts, and schedule messages for later. It also handles reactions, file uploads and downloads, channel housekeeping — creating, renaming, topics, invites, archiving — and user lookups, including presence and do-not-disturb status.

Full operation list
  • Auth — slack_test_auth first to confirm the workspace
  • Channel resolution — slack_find_channels before any channel operation (never pass raw names)
  • Reading — slack_fetch_history (channel, no threads), slack_fetch_thread (by parent ts), slack_search_messages (workspace-wide, supports in:#channel, from:@user, before:/after:YYYY-MM-DD)
  • Sending — slack_send_message with markdown_text for rich formatting; DMs need slack_find_usersslack_open_dmslack_send_message
  • Edit / delete — slack_update_message, slack_delete_message
  • Schedule — slack_schedule_message (Unix timestamp, UTC), slack_list_scheduled, slack_delete_scheduled
  • Reactions — slack_add_reaction, slack_remove_reaction, slack_get_reactions
  • Files — slack_upload_file (text snippet or session file), slack_list_files, slack_download_file, slack_make_file_public, slack_delete_file
  • Channel management — slack_create_channel, slack_rename_channel, slack_set_topic, slack_set_purpose, slack_join_channel, slack_leave_channel, slack_archive_channel, slack_unarchive_channel, slack_invite_users, slack_remove_user, slack_list_members
  • Users — slack_find_users, slack_get_user_info, slack_list_users, slack_get_user_profile, slack_set_user_profile, slack_get_user_presence, slack_get_dnd_status

Try asking"summarise the last 3 days of #product activity in Slack and send the summary to the channel."

My Zoom (14 tools via MCP)

What it does — manages Zoom meetings and recordings on your connected account. It can list, schedule, update, and cancel meetings (timezone-aware), pull participant lists from past meetings, and browse and download cloud recordings. It also fetches AI Companion meeting summaries, reads your webinars, and searches your Zoom contacts.

Full operation list
  • Account — zoom_get_user(userId='me') to confirm account and license
  • Meetings — zoom_list_meetings (pass type='upcoming'), zoom_get_meeting, zoom_create_meeting (always include timezone to avoid shifts), zoom_update_meeting, zoom_delete_meeting
  • Participants — zoom_get_past_participants (ended meetings only, paid accounts only)
  • Recordings — zoom_list_recordings (max 1-month date range), zoom_get_recording (download URLs), zoom_delete_recording (trash is recoverable, delete is permanent)
  • Summaries — zoom_get_meeting_summary requires a paid Zoom plan with AI Companion enabled
  • Webinars — zoom_list_webinars, zoom_get_webinar (Pro+ with Webinar add-on)
  • Contacts — zoom_search_contacts

Heads up — many features (recordings, past participants, AI summaries, webinars) require a paid Zoom plan.

Try asking"schedule a 45-minute Zoom meeting titled 'Board prep' next Tuesday 3pm UK time with these attendees: …"

My Facebook (34 tools via MCP)

What it does — Facebook Page management. It reads page posts with engagement counts, pulls page insights, follower metrics, and brand mentions, and publishes text, photo, and video posts — always confirming with you before anything goes public. It can edit or delete posts, manage comments and reactions, work with scheduled posts, and read and reply to the Page's Messenger conversations.

Full operation list

Always start with fb_list_pages to discover valid page IDs.

  • Read posts — prefer fb_get_page_posts over fb_get_post (doesn't need pages_read_engagement permission). Pass rich fields to get engagement counts in one call.
  • Page info — fb_get_page_details, fb_get_page_insights (metrics: page_follows, page_media_view, page_post_engagements, page_video_views), fb_get_page_roles, fb_update_page_settings
  • Tagged / mentions — fb_get_tagged_posts for brand monitoring
  • Publish — fb_create_post (text/link), fb_create_photo_post (public image URL), fb_create_video_post (direct MP4 URL). Confirm before publishing.
  • Edit / delete — fb_update_post, fb_delete_post (irreversible — confirm)
  • Engagement — fb_get_post_reactions, fb_create_comment, fb_update_comment (edit or hide/unhide), fb_delete_comment, fb_like, fb_unlike
  • Schedule — fb_get_scheduled_posts, fb_publish_scheduled_post, fb_reschedule_post
  • Messenger — fb_get_conversations, fb_get_conversation_messages, fb_send_message (text within 24h window or message tag), fb_send_media_message

Post IDs are in pageId_postId format.

Try asking"pull last week's page insights and summarise which posts drove the most engagement."

My LinkedIn (8 tools via MCP)

What it does — publishes posts and comments on connected LinkedIn accounts. Different from LinkedIn Profile Lookup (which reads public profiles). It can write a post under your name, attach images generated in the session, comment on a post, and delete a post it published.

Full operation list
  • Start with linkedin_get_my_info to confirm the account and get the person URN (needed for posting).
  • linkedin_create_post — post with the author URN (urn:li:person:{id}). For images, pass session file paths in the images array (backend handles upload).
  • linkedin_create_comment — comment on a post URN.
  • linkedin_delete_post — remove a post by URN.

Try asking"write a LinkedIn post announcing our new feature — professional tone, 150 words, include the image I generated in out/banner.png."

My Twitter/X (27 tools via MCP)

What it does — reads, posts, engages, and pulls analytics on connected X (Twitter) accounts. Different from Twitter/X Search (which queries the public firehose without auth). It can look up users and tweets, search the last 7 days, read your home timeline, post tweets and threads with images or video, like, retweet, bookmark, and follow, and pull impressions and engagement analytics on tweets you own. Up to 3 handles per workspace. Every write action is safety-gated under the Twitter (X) card in the Safety Center.

Full operation list
  • Identitytw_get_me (runs automatically on connect, labels the account @yourhandle).
  • User lookuptw_lookup_user, tw_lookup_users, tw_get_user_by_id for profile details, follower/following counts.
  • Tweet lookup & searchtw_lookup_tweet, tw_lookup_tweets, tw_recent_search (last 7 days, operator syntax), tw_home_timeline, tw_get_quotes, tw_get_retweeters.
  • Publishtw_post_tweet (text, reply, quote, with media), tw_delete_tweet (irreversible — always asks).
  • Engagetw_like / tw_unlike, tw_retweet / tw_unretweet, tw_bookmark / tw_unbookmark, tw_follow_user / tw_unfollow_user, tw_get_bookmarks, tw_get_liked_posts.
  • Mediatw_upload_media (simple path for images/GIFs) plus chunked video: tw_init_media_uploadtw_append_media_uploadtw_get_media_status. Attach returned media_id when calling tw_post_tweet.
  • Analyticstw_get_post_analytics returns impressions, engagements, video views, profile clicks for tweets you own.

Heads up — connecting is a one-click OAuth login; no developer portal, no API keys. Write actions are safety-gated by default (Ask) under the Twitter (X) card in the Safety Center. If posting suddenly fails, the platform-side X API pool may need a refill — flag it in Discord.

Try asking"draft a 5-tweet thread announcing the new release, attach out/launch-hero.png to the first tweet, and post from @yourhandle."

My Discord (10 tools via MCP)

Release note: Discord — Identity, Servers & Invites.

What it does — read-only account surface for Discord, OAuth2 user-scoped. Cannot send or read channel messages, list channels, or create them — that's a Composio toolkit limitation, not a config gap. Use Slack / Teams / Gmail for community messaging. Discord is for membership audits (your servers, roles, and permissions, with member and online counts), invite resolution, profile housekeeping (username and avatar changes, leaving a server — both confirmed first), and an inventory of the Twitch / Spotify / Steam / YouTube accounts linked to your Discord.

Full operation list
  • Identitydiscord_get_my_user (username, global display name, user ID, locale, email if scoped). Run this first on every new session to confirm the connected account.
  • Servers (guilds)discord_list_my_guilds (paginated via before/after/limit, max 200 per call; optional with_counts for member + online counts), discord_get_my_guild_member (your roles, nickname, join date, permissions inside one guild).
  • Invitesdiscord_resolve_invite accepts plain codes (abc123), vanity codes (discord-api), or full discord.gg/... URLs. Optional member/presence counts and a scheduled event ID.
  • Public guild metadatadiscord_get_guild_widget (widget JSON for guilds that have it enabled), discord_get_guild_template (template details by code from discord.new/{code}).
  • Profile managementdiscord_modify_my_profile (change username and/or avatar, or remove avatar; Discord caps username changes at 2 per hour; publicly visible — Alfrada OS confirms first), discord_leave_guild (irreversible — Alfrada OS confirms first).
  • Connections & cataloguediscord_list_my_connections (Twitch / Spotify / Steam / YouTube accounts linked to your Discord — requires the connections OAuth scope), discord_list_sticker_packs (Discord's Nitro sticker catalogue).

Heads up — connecting is a one-click OAuth login. email and connections scopes are opt-in at connect time; if you skipped them, the tools that depend on them return blanks. The widget tool only works on guilds where the admin enabled the widget. Up to 3 Discord accounts per workspace — name the account (display name or email) when you mean a specific secondary one.

Try asking"list my Discord servers with member and online counts, sorted by size" or "resolve discord.gg/python and tell me what guild it's for".

My Canva (14 tools via MCP)

Release note: Canva — Create, Import, Export & Organise Designs From Chat.

What it does — design creation, file import/export, and asset management on a connected Canva account. It creates new designs (docs, presentations, whiteboards, or custom pixel sizes), imports PDF / DOCX / PPTX / XLSX / PSD / AI files as editable designs, and exports designs to PDF, PNG, JPG, PPTX, GIF, or MP4. It can also upload images, videos, audio, PDFs, and fonts as reusable assets, and browse or search your designs and project folders. Imports, exports, and uploads run as background jobs — Alfrada OS waits on them and surfaces download URLs, saved session files, and design IDs inline.

Full operation list
  • Identitycanva_get_user returns the connected account's display name. Run first to confirm which Canva you're acting on.
  • Createcanva_create_design takes a single design_type arg: a preset ('doc', 'presentation', 'whiteboard') or a custom object like {"type": "custom", "width": 1080, "height": 1080} (1–8000 px per side). Optional title (1–255 chars) and optional asset_id to seed with a freshly uploaded image.
  • Import (PDF / DOCX / PPTX / XLSX / PSD / AI → editable design)canva_import_design with title and attachments: ["in/deck.pdf"] (one file per call). Async — poll canva_get_import_status with jobId until status is success.
  • Export (PDF / PNG / JPG / PPTX / GIF / MP4) — call canva_get_export_formats with designId first to see supported formats, then canva_export_design with design_id and format as an object: {"type":"pdf"}, {"type":"png","size":{"width":1080,"height":1080}}, {"type":"jpg","quality":75}, {"type":"pptx"}, {"type":"gif"}, or {"type":"mp4","quality":"horizontal_1080p"}. Async — poll canva_get_export_status with exportId; successful exports auto-save to session in/ when possible, otherwise signed download URLs (valid ~24h).
  • Upload assetscanva_upload_asset accepts images / videos / audio / PDFs / fonts from session paths (attachments: ["in/logo.png"]). For public HTTPS URLs, use canva_upload_asset_from_url with url + name. Async — poll canva_get_upload_status with jobId for the assetId.
  • Readcanva_get_design (title, owner, thumbnail, edit URL, page count), canva_get_asset (name, tags, MIME type, thumbnail).
  • Browsecanva_list_designs (text search, ownership filter any/owned/shared, sort orders, continuation-token pagination), canva_list_folder_items (use folderId: "root" for the top-level project folder; optional item_types: ["design","folder","image"]).

Heads up — connecting is a one-click OAuth login; no developer portal, no API keys. Export URLs expire after ~24 hours — save the file or paste it where you need it that day. Up to 3 Canva accounts per workspace — name the account (the Canva display name shown in Integrations) when targeting a specific secondary one. Paid Canva features (MP4 export, brand kit, Pro templates) still require a paid Canva plan; Alfrada OS doesn't proxy those entitlements.

Try asking"import in/strategy-draft.pptx as a Canva design called 'Board April 2026', then export it back to PDF when it's ready" or "list my 10 most recently modified Canva designs with thumbnails".

My Hunter (8 tools via MCP)

Release note: Hunter — Find, Verify & Enrich Professional Contacts.

What it does — professional email discovery, verification, and B2B enrichment via Hunter.io on a connected account. It finds the most likely work email from a name and company, lists the public emails on a domain, and verifies whether an address is deliverable. It also enriches people (email or LinkedIn handle in — name, title, socials out) and companies (domain in — industry, size, location out), and discovers companies in Hunter's B2B dataset by industry, headcount, location, or a plain-language query. Discovery and enrichment only — there is no leads CRM.

Full operation list
  • Email finderhunter_email_finder resolves the most likely work address from a full name + domain/company. Treat accept_all / risky as low confidence.
  • Domain searchhunter_domain_search lists public emails for a bare domain (acme.com, never https://www…). Higher limits need a paid plan.
  • Verifier & volumehunter_email_verifier checks deliverability for one address; hunter_email_count returns free volume stats by domain (no credit cost — use before bulk searches).
  • Enrichmenthunter_people_enrichment (email or LinkedIn handle → name/title/socials), hunter_company_enrichment (domain → industry/size/location), hunter_combined_enrichment (person + company in one call).
  • Company discoveryhunter_discover_companies filters Hunter’s B2B dataset by industry, headcount, location, or a natural-language query. Many filters are Premium-only.

Heads up — Alfrada OS won't invent email addresses: if Hunter misses, it says so. A shared monthly quota and Hunter plan caps apply. Up to 3 Hunter connections per workspace when multi-account is enabled — name the account when targeting a secondary one.

Try asking"find Jane Doe's work email at acme.com and verify it" or "discover 10 SaaS companies in Germany with 50–200 employees, then enrich the top three domains".


4. Creative tools

Image, video, audio, memes, diagrams — 10 tools. Product notes: Seedream 5.0 Pro Image Lab, xAI Image Lab, Runway video, Medical Vision, Image Consistency.

Medical Vision (experimental)

Release note: Medical Vision — AI Second Opinion On Your Scans.

What it does — AI second opinion on medical images and records using MedGemma 4B, running entirely on your own Ollama runtime (no PHI leaves the deployment). Listed in the catalog under Health & Clinical. Reads chest X-rays, CT / MRI / ultrasound, fundus & retinal scans, dermatology photos, histopathology slides, and clinical documents (lab panels, prescriptions, discharge summaries). Two modes: clinical (default — structured radiology-style markdown with an explicit second-opinion disclaimer) and inspect (forensic re-read of any non-medical image, no clinical framing). Up to 4 images per call; the agent can call the tool again on the same image with a different question to focus on a new region — the model stays warm across calls.

When to reach for it — you uploaded an X-ray or lab report and want a machine-generated second look before your clinician appointment; you've already analysed an image once and want a focused follow-up ("now focus on the spine"); or you want a forensic re-read of any image (dashboard, chart, whiteboard) that's richer than the default upload caption.

Heads up — MedGemma is an experimental research model, not a medical device, not for primary care, and not a substitute for a licensed clinician. Every clinical response carries that disclaimer. If symptoms are severe or urgent, contact emergency services first.

Try asking"here's my chest X-ray — describe the cardiac silhouette, lung fields and anything a radiologist should review".

Image Consistency

Release note: Image Consistency — Same Person, Same Object, Same Style, Verified Locally.

What it does — compares 2–4 session images in one pass and returns a structured verdict: same person, same object, or consistent visual style across them, with a confidence score, a one-line description of each image, and specific differences. Runs on Gemma 4 12B on the local Ollama runtime — images never leave the deployment, no per-call cost. Every comparison writes a markdown report to out/ and indexes it. The same engine automatically audits the video editor's restyle pipeline and attaches an advisory frame-drift verdict to every restyle result.

When to reach for it — you generated several images of the same character and want to verify they stayed consistent; you have two product photos and need to know whether they show the same physical item; you restyled a video and want to know if the look drifted between frames; you want a same-person check across two photos.

Heads up — the verdict is advisory: reliable on clear cases, but treat low-confidence calls (below ~70%) as a prompt to look yourself. Max 4 images per call.

Try asking"compare out/hero_v1.png and out/hero_v2.png — same character? Focus on hair, outfit, and face."

Image Lab

Release note: Image Lab Defaults To Seedream — Gemini For 4K And Batches.

What it does — image generation and editing across ByteDance Seedream 5.0 Pro via OpenRouter (the default renderer), Gemini, and xAI via OpenRouter. Two operations: create (new image) and edit (modify an existing session image — always prefer this when the user references something already in the session). Resolutions: 1K (fast), 2K (recommended), 4K on Gemini; Seedream and xAI support 1K/2K. Ten aspect ratios: 1:1, 2:3, 3:2, 3:4, 4:3, 4:5, 5:4, 9:16, 16:9, 21:9. Gemini can generate 1–10 images per call; Seedream returns one — so a request for 4K or several images in one call is routed to Gemini automatically. Optional Google Search grounding for Gemini brand/product accuracy.

Provider choices:

  • Seedream / auto (bytedance-seed/seedream-5-0-pro) — the default renderer: precise product and identity-preserving edits, lifelike commercial scenes, and multi-reference compositions with up to 14 source images. One 1K/2K output per call via OpenRouter.
  • Gemini (provider: "gemini") — true 4K, up to 10 images in one call, and Google Search grounded image/text outputs. The default route switches to it by itself whenever a call needs 4K or more than one image.
  • xAI (provider: "xai", x-ai/grok-imagine-image-quality) — fast 1K/2K posters, ads, packaging, menus, social graphics, clean typography, named locations, brand variants, and public-figure creative via OpenRouter.
  • Automatic retry on refusal — if the renderer declines a prompt on the default route, the same call is retried once on xAI; the response tells you which provider rendered the final file.

Prompt styling (style_preset) — by default ("auto") a style and color direction is appended to the prompt, inferred from its keywords, and matched to the light/dark UI theme. Force one with "infographic", "presentation", "creative", or "neutral", or pass "none" to send the prompt verbatim. Reach for "none" when the prompt already pins the exact look — precise edits, brand-spec renders, photography — so the inferred style cannot override it.

When to reach for it — illustrations, concept art, marketing visuals, infographic assets, cover images, product mockups, ad posters, and editing an existing image to change a detail. The default already gives you Seedream's controlled commercial edits and multi-reference compositions; use xAI when text rendering or named-entity fidelity matters, and Gemini for 4K or several variations at once. Use Pexels instead when you need licensed real-world stock media.

Heads up — Seedream returns one image per call. xAI's ability to render brands, public figures, places, or other named entities is a model capability, not legal clearance. For third-party franchises/characters, use descriptive homage language unless you own or have rights to the IP.

Try asking"use Seedream to edit out/product.jpg — preserve the exact product, logo, reflections, and camera angle; change only the set to a warm luxury hotel at dusk." — or "design a 16:9 launch poster for Alfrada OS with a bold platinum headline — dark graphite background, soft-gold accents, cinematic lighting."

Pexels Stock

What it does — searches Pexels' royalty-free library of professional stock photos and videos and saves the best matches straight into your session. You can browse curated photo picks and popular videos, filter by orientation (landscape, portrait, square), colour, and clip length, and choose quality up to 4K for video. Every saved file arrives with the photographer's name and a ready-to-use credit line ("Photo by X on Pexels"), and is immediately usable by Presentation Create, the Video Editor, and Image Lab in the same turn.

When to reach for it — when a deliverable needs real people, real places, or real products rather than AI renders: stock slides, authentic social imagery, B-roll and background footage, hero images. Use Image Lab for concept art and mockups; use Google Images for a specific real-world subject from the open web with licence filtering.

Try asking"find three landscape stock photos of a busy Tokyo crossing and add them to my deck with credits" or "get me 10 seconds of HD ocean-wave b-roll for the intro."

Video Lab (beta)

Release notes: Video Lab Works On Your Footage — YouTube Rips, Aleph 2 Edits, Magnific 4K, Act-Two · Grok Imagine Video 1.5 — Native 1080p + Synchronized Audio · Seedance 2.5 First — Retention Tags + 30s Clips · MiniMax Hailuo 3 Joins Video Lab — Native-Audio 2K, And It Goes First · xAI Imagine Video — Native Audio + Fallback Chain · Runway Gen-4.5 Takes Over Video Generation.

What it does — short AI-generated video across four providers, picked by the provider knob:

  • Seedance 2.5 (provider: "seedance") — ByteDance via OpenRouter. Native-audio text-to-video and image-to-video with first+last frame, 4–30 seconds, 480p/720p. Zero prompt retention on OpenRouter. First choice on provider: "auto". Distinct from Runway seedance2.
  • xAI Imagine Video (provider: "xai") — Grok Imagine Video via OpenRouter. Text-to-video, image-to-video, and reference-to-video. Native audio. 1–15 seconds, 480p / 720p. Prompts retained 30 days; not trained on.
  • MiniMax Hailuo 3 (provider: "minimax") — native-audio text-to-video, image-to-video, and reference-to-video (up to 9 stills) at 2K, 5–15 seconds. Toggle audio with minimax_audio. Prompts retained (duration unspecified); not trained on.
  • Runway (provider: "runway") — Gen-4.5 / Gen-4 Turbo / gen3a_turbo and the Veo / Runway seedance2 audio-native models. Best for first→last keyframe morphs and explicit model control. 2–10 seconds per clip. Prompts retained (duration unspecified); not trained on.

provider: "auto" (the default) runs Seedance → xAI → MiniMax → Runway and each call's response carries a fallback_chain block plus a data_policy tag. All video jobs briefly keep the rendered MP4 for download (ZDR cannot apply).

Operations: create (text-to-video), create_with_image (animate a still; add last_frame_image for first→last keyframe animation on Seedance / MiniMax / Runway gen3a_turbo / veo3.1 / veo3.1_fast / seedance2), edit_video (prompt-edit existing footage ≤30s via Runway Aleph 2), upscale_video (Magnific creative upscale to 720p–4K), character_performance (Runway Act-Two motion transfer: a character still performs an acting video), and enhance_prompt (rewrite only, no video — local Ollama, not a render charge). MiniMax also accepts reference_images (max 9 session filenames); Seedance takes up to 30 reference_images plus 10 reference_videos for scene continuation / motion transfer. Optional during generation: watermark_text or watermark_image, applied by the NVENC-accelerated sidecar. Audio-native renders are checked automatically for a real (non-silent) soundtrack and come back with a storyboard summary; if a generated clip already has sound, Alfrada OS checks with you before layering narration over it.

When to reach for it — product demos, social clips, marketing shorts, animated stills, hero loops, B-roll, short dialogue scenes, longer 20–30s Seedance shots, and reference-to-video character consistency. Longer than 30s: chain clips in the Video Editor, or use the Polished narrated video playbook.

Try asking"make me a 12-second café scene with sound — a barista calling out an order" — or "animate this product still with a slow 20-second orbit."

CUDA Video Editor (beta)

Release notes: Video Editor — Reframe, Jump-Cut, Captions, Overlays, And Animated Slides · Live Video Studio

Full guide: CUDA Video Editor

What it does — production-quality video editing (clips, voice-overs, music, multi-resolution export) on the Strategize CUDA cluster. While Alfrada OS runs this tool (or Video Lab), open Live → Video Studio in the Work panel: one session-scoped timeline editor with a step list during the build and preview + drag-trim-reorder after the finishing pass.

Quick path — the everyday flow: Alfrada OS ingests your video (AI scene breakdown + transcription), edits a timeline draft (cuts, subtitles, transitions, effects, audio, slide cards), builds it into the finished file, and can report the timeline's status at any point. Describe the result you want — you never need to name the internal steps.

What it can do:

  • Inspect — read a video's duration, resolution, frame rate, codecs, and whether it has audio. Ingest additionally transcribes the audio with word-level timing, labels speakers, and lifts any embedded subtitle track — which is what makes automatic captions a one-line request.
  • Cut to length — hard-trim a video to an exact duration ("cut it to 30 seconds").
  • Extract and reformat — pull a segment out of a longer video, change its speed, and resize it for a new format (letterbox, crop-to-fill, blurred background, stretch). Sizes can be exact pixels, a platform preset (TikTok, Reel, YouTube Short, 2K cinema…), or "keep the source size" so 4K stays 4K. The full guide is the canonical list of resize modes.
  • Reframe and transform — crop to a target aspect without letterboxing, rotate or flip sideways phone footage, reverse, loop, fade in, change frame rate (with true interpolated slow motion), stabilise, denoise, sharpen, or deinterlace — all applied in a single render.
  • Overlays — picture-in-picture, side-by-side or stacked comparison, timed image overlays, text cards, and lower-third name badges whose text wraps to fit.
  • Stitch with transitions — join clips with fades, wipes, zooms, and more; the full guide is the canonical transition list.
  • Subtitles — burn styled subtitles (font, size, position, background opacity) into the video, export them as an .srt/.vtt sidecar, mux them in as a switchable track, or highlight them word-by-word karaoke-style from the transcript's word timings.
  • Audio surgery — layer narration and music onto a video, merge audio files into one track, split audio into short segments, extract a video's soundtrack to MP3/WAV/AAC/FLAC, remove all baked-in audio before replacing it, or speed narration up / slow it down without re-recording.
  • Colour grades — list and apply one-shot colour looks: 12 presets — cinematic, black-and-white, matrix green, optimus blue, SpaceX black, Instagram warm, Instagram cool, vintage 8mm, cyberpunk, documentary, noir, and sketch — plus custom grades built directly when none of the twelve fit.
  • Slides and explainers — render title cards, bullet slides, image cards, and Manim-driven animated slides (including typeset equations) as standalone clips to stitch between live footage. Image slides can carry a Ken Burns move instead of sitting still.
  • Scenes and jump cuts — detect shot boundaries (and optionally cut one clip per shot), or strip the dead air out of a talking-head recording and report how much came off.
  • Watermark — stamp text or a logo image onto a video.
  • GIF output — the build step can compile the timeline to an optimised-palette GIF instead of MP4.
  • Restyle pipeline — "make this trailer Ghibli-style": the editor restyles the ingested video's deduped key-frames one by one via Image Lab's identity-preserving edit, reassembles them into video, and overlays the original audio. A single bad frame can be redone alone at its timestamp, and every restyle result automatically gets an advisory frame-drift verdict from the local Image Consistency engine.
  • Multiple drafts — the editor keeps multiple timeline drafts ("canvases") per session: create or fork a draft, switch between them, ask for AI suggestions, review a change log, and rebuild any draft's output.
  • Finish — a mandatory finishing pass adds the closing fade and brand watermark and marks the file as the deliverable, which unlocks the Live Video Studio timeline.
Full operation list

All 37 operations, exactly as the tool defines them:

  • Canvas flow — ingest (upload + AI-analyse a video), edit (mutate the timeline canvas), build (compile canvas to MP4/GIF), status (session state + canvas), rebuild (recompile a canvas output)
  • Drafts — create_canvas (new or forked canvas), list_canvases, set_active_canvas, get_canvas, update_canvas (apply a merge patch), suggest_canvas (AI canvas ideas), get_change_log (canvas change entries)
  • Precision cuts — probe (duration, per-stream video/audio durations, resolution, has_audio, codecs), extract_clip (cut a segment, optional resize/speed), trim (an exact duration, optionally from a start offset), pad_video (add freeze-frame or black at the head/tail of a clip), concat (join segments with transitions; transition_duration: 0 is a hard cut, gap_before inserts black + silence between scenes), split_scenes (detect shot boundaries, optionally emitting one clip per shot), remove_silences (jump-cut the dead air out of a talking head)
  • Picture — transform (crop/reframe by aspect, rotate, flip, speed, reverse, loop, fade in/out, fps with interpolation, stabilise, denoise, sharpen, deinterlace, or a raw allowlisted filter chain — all in one re-encode), overlay (image, video, text card, or lower-third; PiP, side-by-side, stacked, or timed)
  • Audio — mix_audio (layer audio tracks onto a video; per-track start/end timeline offsets; output sound is always locked to the picture length), strip_audio (remove all audio streams), extract_audio (lift a video's soundtrack out to mp3/wav/aac/flac), mix_audio_only (layer audio files into one MP3, or sequence: true to play them one after another with gap_seconds of silence), speed_audio (change audio speed), split_audio (chunk audio into 0.5–10s segments)
  • Look — list_filters / apply_filter (colour-grade presets, plus a custom filter chain), watermark (stamp text/image onto video), burn_subtitles (burn subtitles into video, optionally word-by-word), export_subtitles (write cues to an .srt/.vtt sidecar), render_slide (template or Manim slide → MP4, with optional Ken Burns motion on stills)
  • Restyle — restyle_frames (restyle ingested key-frames), compose_from_frames (restyled frames → silent MP4), restyle_video (one-shot restyle + compose + audio)
  • finalize — mandatory final fade-out on every delivered video (optionally with an opening fade and a muxed-in switchable subtitle track)

Live Video Studio (Work panel): Phase 1 step list while tools run → Phase 2 preview + horizontal timeline after the finishing pass. User can reorder clips, trim, set transitions/colour grade, move audio, and Update preview to rebuild without re-prompting.

When to reach for it — trimming, transitions, highlight reels, subtitles and caption files, audio mixing, colour grading, reframing to vertical, jump-cutting a talking head, overlays and lower-thirds, full-video restyles, polished narrated shorts. Pairs with Video Lab for long-form narrated video.

Heads up — editing operations re-encode the whole file, so they scale with the source's length. Long renders move to a background job and report back in the same conversation rather than freezing the turn.

Try asking"ingest in/demo.mp4, trim to the 90 seconds where I show the integrations tab, add captions, finalize, then I'll tweak the timeline in Live Video Studio." — or "restyle this trailer in a watercolor style, but keep the faces and on-screen text."

Music Lab

What it does — original music, vocals, and instrumentals powered by Suno (models V3_5 / V4_5 / V5). Supports vocal tracks (male / female) or instrumental only. Two authoring modes: Idea mode (all guidance in a single prompt — recommended) or Structured (explicit style + title + lyrics).

When to reach for it — background music, jingles, podcast intros, soundtracks for video, demos for songwriting.

Try asking"make a 90-second upbeat instrumental in the style of electronic lo-fi, no vocals, for a product demo."

Meme Tool

What it does — creates memes from 40+ classic templates with custom captions. Multi-box text support. Ask for the template list to see them all.

When to reach for it — social media content, team chat moments, light-touch internal comms.

Try asking"make a Distracted Boyfriend meme: me, my current stack, the new AI framework everyone is hyping."

Mermaid Diagram

What it does — renders Mermaid diagrams inline: flowcharts, sequence diagrams, Gantt charts, org charts, ER diagrams, state diagrams. Downloads as SVG/PNG from the diagram card.

When to reach for it — visualising processes, system architecture, timelines, org structures, decision trees.

Try asking"draw a sequence diagram for the user signup flow: browser → API → Stripe webhook → database → email."

Cats (18 tools via MCP)

What it does — browse cat images, breeds, and facts via TheCatAPI. It can search cat photos (up to 25 per call, filterable by size, format, or breed), pull machine-generated labels for an image, and look up breed profiles. It also browses themed categories (hats, sunglasses, boxes and more) and manages your favourites, votes, and uploaded images.

Full operation list
  • Images — cats_search_images (up to 25 per call, filter by size / mime type / breed-tagged), cats_get_image, cats_get_image_breeds, cats_get_image_analysis (ML labels)
  • Breeds — cats_search_breeds, cats_get_breeds (paginated), cats_get_breed (by ID like pers, beng, siam)
  • Categories — cats_list_categories (hats, sunglasses, boxes, etc.)
  • Favourites — cats_create_favourite, cats_list_favourites, cats_get_favourite, cats_delete_favourite
  • Votes — cats_create_vote (1=up, 0=down), cats_list_votes, cats_get_vote, cats_delete_vote
  • Uploads — cats_list_uploaded_images, cats_delete_image

Try asking"show me 5 Bengal cat images and tell me about the breed."


5. Documents & data

Markdown, reports, Excel, in-place Word / PPTX / PDF edits, Google & Microsoft 365 (MCP), scheduling — 20 tools.

Markdown Editor

What it does — creates, edits, reads, lists, and deletes markdown files. Exports as .md, .html, or .doc from the Markdown tab download menu.

When to reach for it — drafts, long-form writing, notes, anything you want to edit inline in the Work Panel.

Try asking"draft a 1,200-word thought-leadership post on RAG vs fine-tuning, save it as out/post.md."

Report Generator

What it does — generates polished deliverables. Four output formats: PDF (traditional reports), DOCX (Word documents), Rich Slide Deck (professional slides — the default), or All (all three at once). Supports charts, tables, headers, embedded images, executive summaries.

When to reach for it — final deliverables after research is complete. Board reports, client briefings, research digests. Best as the last step of a workflow, not the first.

Try asking"take the research we just did and produce a 10-page PDF report with an executive summary, 3 charts, and a one-page conclusion."

Presentation Create

What it does — builds brand-new multi-slide presentations with structured slide templates, layouts, text, images, and themes. The create schema requires a real non-empty slides array, and every slide must include slide_type and content. Exports to PowerPoint from the Presentations panel. For Google Slides output, use the Google Slides MCP integration below.

When to reach for it — pitch decks, board presentations, training materials, executive summaries in slide format.

Try asking"build a 10-slide pitch deck for my consumer AI startup: problem, solution, market, traction, team, ask."

Presentation Edit

What it does — edits an Alfrada OS-native deck that was already created by Presentation Create. It requires the deck's presentation_id and a focused edits array, so the agent can update, insert, or delete specific slides without rebuilding unchanged slides.

When to reach for it — tighten slide 4 of a generated deck, append a new appendix slide, replace one KPI slide, or delete a slide from a deck Alfrada OS just created. For uploaded .pptx files, use PPTX Patch instead.

Try asking"update the deck you just created: replace slide 6 with a clearer pricing slide, and add one appendix slide with the assumptions."

PPTX Patch

Release notes: Edit Uploaded PowerPoint Decks In Place, Design Intact · Uploaded Decks And PDFs — Charts, Forms, Pictures, Visual Checks.

What it does — surgically edits a .pptx you uploaded instead of rebuilding it. Inspect returns every slide's shapes (shape_id), text (paragraph and soft line breaks), table cells, native category charts, pictures — including filled picture placeholders, the usual home of a template hero photo — speaker notes, animation/transition XML, and a rendered visual read of the first 10 slides. Patch then applies narrow ops: replace text (real line breaks, not leftover _x000B_), recolor text / fills / outlines, set font size, wrap, and style, edit a table cell by row and column, replace bar/column/line/pie/doughnut/area chart data and titles, swap an embedded picture or placeholder photo, contain-fit a new picture into a box, move / resize / hide / delete a shape, duplicate / delete / reorder slides, update notes, and set core properties (so a client deck does not ship with the template vendor as author). Every patch runs a visual check on changed slides (up to 60; the rest are listed as unchecked) plus a text-integrity scan, and reports what changed, what stayed identical, and anything it could not guarantee. The original design, fonts, master slides, and existing animations are preserved.

When to reach for it — "redo slide 3 but keep the design", swapping a logo or template hero photo, updating a native chart, filling a branded template, fixing a typo or a status-chip colour, deleting or reordering slides. Use Presentation Create for a brand-new deck, and Presentation Edit only for Alfrada OS-native decks that already have a presentation_id. New animation authoring is not supported. Linked (non-embedded) images and XY / scatter / bubble charts error rather than silently degrade. Encrypted files that will not open are refused.

Try asking"I uploaded board-deck.pptx — replace the cover image with the new logo, put these figures into the chart on slide 4, and fix the title on slide 2. Don't change anything else."

Doc Patch

Release notes: Edit Uploaded Word Documents In Place, Formatting Intact · Workbooks And Word Files Now Fill Like The File You Sent.

What it does — surgically edits a Word .docx you uploaded or a saved template instead of rebuilding it. Inspect returns every paragraph and table with stable references (para:i, table:i, image:i), embedded images, comments, headers/footers, styles in use, and every left in the file. A saved templates/<file>.docx seeds a working copy in out/ — the stored template is never mutated — and attaches a rendered visual read of the pages so fills are mapped before the first edit. Patch then applies narrow run-aware changes — replace exact text while keeping the run's formatting, replace a paragraph, edit/add/insert/delete table rows and cells, swap an image, insert or delete blocks — against a required preservation contract. A visual check (slow; for filled templates and layout-risky docs) renders the file and flags overflow, leftover , garbled codes, and broken images; a text-integrity scan still runs if rendering is unavailable. It can also create a brand-new .docx from structured sections, including an embedded session image (aspect preserved, capped to the 6.5in page body; a missing file is reported without sinking the document). Surrounding text, run-level formatting, and table styles are preserved.

When to reach for it — filling a branded pack you saved under Settings → Templates, "change the hour total in section 2.2.1 but keep the rest", updating a table in the Word file you uploaded, "send it back as Word". Use Report Generator (output_format='DOCX') instead for a brand-new report authored from markdown, and Markdown Editor for notes and drafts.

Try asking"I saved agreement.docx as a template. Fill every leftover placeholder for Ooredoo, keep the letterhead, and tell me if any placeholder is still showing."

PDF Patch

Release note: Uploaded Decks And PDFs — Charts, Forms, Pictures, Visual Checks.

What it does — surgically patches a .pdf you uploaded instead of re-authoring it. Inspect and read-page attach a rendered visual analysis (first 10 pages) so formatting claims come from the page, not from blind extracted text. Locate returns exact text coordinates in PDF points. Patch then fills AcroForm fields, sets title/author metadata, deletes / rotates / reorders / merges pages, stamps text or image overlays (x/y or a page-corner anchor), watermarks, strips annotations, or fixes a small text defect by covering the exact glyph rect and restamping (a typo, a wrong-weight letter, a date — visual only, not redaction). Extract saves a page range as a new PDF. Every patch runs a visual check on touched pages (up to 40). Untouched content streams stay as they were.

When to reach for it — "fill this form", "delete page 3", "merge the appendix", "watermark it Confidential", "the H on page 1 is the wrong weight". Use Report Generator for a brand-new PDF. A wholesale paragraph rewrite cannot reflow in place — convert those pages to Word with Doc Patch and say the layout is reconstructed. If inspect reports a filled signature field, a patch invalidates that signature — Alfrada OS should warn first. Overlay and restamp text is Latin-1 only.

Try asking"I uploaded application.pdf — fill the name and address from these notes, watermark every page 'Draft', and don't rewrite the rest of the file."

HTML Dashboard

Release note: HTML Dashboards — Beautiful, Interactive Reports That Render Live And Export To PDF.

What it does — builds beautiful, self-contained interactive HTML dashboards (hover tooltips, live charts, KPI cards) from six prebaked templates — executive_summary, kpi_grid, research_report, comparison, timeline, blank. Renders live in the Dashboards panel, refreshes the same dashboard in place when data changes, and exports to PDF, PNG, or raw HTML. On-brand with automatic dark mode.

When to reach for it — live KPI walls, operating dashboards, interactive research readouts, head-to-head comparisons, roadmaps — anything that benefits from interactivity rather than static slides. Use Presentation Create for new .pptx decks, Presentation Edit for Alfrada OS-native deck revisions, and Report Generator for static PDF/DOCX prose.

Try asking"build an interactive executive dashboard of these Q2 metrics with a trend chart, then let me export it as a PDF."

Excel Ops

Release note: Workbooks And Word Files Now Fill Like The File You Sent.

What it does — creates and edits real .xlsx workbooks. A saved template under Settings → Templates wins: inspect templates/<file>.xlsx first, which seeds a working copy in out/ (the stored file is never edited), returns , cell refs, and style facts, and attaches a visual read of the rendered sheets. Tokens are filled with find-and-replace; leftover is reported. Numeric strings become numbers and formula text becomes live Excel formulas (a bare MAX(...) is rescued; 007 stays text). A visual check (slow; for template fills and layout-risky sheets) flags clipped #### columns, spilling text, overlaps, and leftover tokens. Step-by-step editing also covers worksheets, charts, tables, freeze panes, dropdowns, auto-fit columns (widths 8–60 characters; longer text wraps), and session images (webp/TIFF convert to PNG; unsized images cap at 960×720).

When no template matches, six presets style a new workbook — you still supply every question, criterion, or line item:

  • questionnaire — assessment forms with sections and scored items
  • comparison_matrix — vendor/product comparisons with criteria
  • executive_report — multi-section reports with headers and rows
  • tracker — project/action trackers with typed columns and validation lists
  • financial_model — assumptions, periods, line items
  • dashboard — KPI sections with named KPIs

When to reach for it — filling the branded model you already saved, patching an uploaded workbook without rebuilding it, questionnaires, RFPs, vendor scorecards, trackers, financial models, executive dashboards.

Try asking"Use our saved financial-model template. Fill it for Meridian with these assumptions, don't restyle the sheets, and check the layout before you hand it over."

Task Scheduler

What it does — schedules tasks to run at a set time or on a recurring cadence (daily, weekly, monthly, custom cron). Each run is a separate session with its own objective and deliverable.

When to reach for it — daily morning briefings, weekly competitor monitoring, monthly data pulls, clock-anchored jobs ("every Monday at 8am"). Pairs naturally with Argus (on Pro Max and Black) for autonomous execution.

Try asking"every Monday at 8am, research what's new in AI regulation and email me a summary."

Wake Me

Release note: Wake Me — Alfrada OS Checks Back In The Same Conversation.

What it does — schedules a wake-up that resumes this same conversation after 5 minutes to 24 hours. Alfrada OS re-enters the thread, re-checks what you were waiting on (email reply, long render, price level, approval), and either delivers, re-arms with a longer interval, or stops. Actions: set (arm or replace), cancel, status. One pending wake per conversation; 10 consecutive self-wake turns max (a live reply from you resets the chain); five pending wakes per user across all conversations. Always on — no Tool Selector toggle. Pending wakes appear in Settings → Automations and can be paused there.

When to reach for it — short-horizon follow-through inside the chat you already have: waiting on someone else, waiting for a job to finish, checking back before a deadline. Different from Task Scheduler (separate recurring runs) and unavailable in WhatsApp / Argus chats (Argus heartbeat owns timing there).

Try asking"finish the vendor summary once Alex replies — check back in 45 minutes" or "the export should be done in 20 minutes; wake this chat and verify the file, then add subtitles."

My Google Calendar (28 tools via MCP)

What it does — manages calendar events and availability on your connected Google account. It can list your calendars, read events, and schedule new ones — structured with attendees and times, or created from a plain sentence. It also modifies existing events and finds free slots that work for you and your attendees.

Full operation list
  • gcal_list_calendars — discover your calendars
  • gcal_events_list — read events
  • gcal_create_event — schedule (use this for structured events)
  • gcal_quick_add — simple natural-language event creation
  • gcal_patch_event — modify existing events
  • gcal_find_free_slots — availability lookup

Attendee fields take JSON arrays of emails (["alice@example.com"]).

Try asking"find me a 30-minute slot next week that works for me and alice@example.com, then schedule a meeting titled 'Pricing sync'."

My Gmail (22 tools via MCP)

What it does — full Gmail access. It searches and lists your mail, reads a specific email or an entire thread, sends, drafts, and replies in-thread, and manages labels. It downloads attachments into the session and can attach a session file when sending — one attachment per email (a provider limit).

Full operation list
  • gmail_fetch_emails — search/list
  • gmail_get_message / gmail_get_thread — read a specific email or full thread
  • gmail_send_email — send
  • gmail_create_draft — draft
  • gmail_reply_to_thread — reply in-thread
  • gmail_list_labels / gmail_create_label — label management

Attachments: for downloading, call gmail_get_message first to see attachment IDs, then gmail_get_attachment per file (saved to session in/). For sending, pass session file paths in the attachments array. Only one attachment per email (provider limit).

Try asking"find the latest invoice from Stripe in my inbox, download the PDF attachment, and summarise the charges."

My GDrive (21 tools via MCP)

What it does — full Drive file operations. It finds files and folders by name, type, date, or content, creates empty Docs, Sheets, and Slides, new folders, or files from text, and copies, moves, trashes, restores, or deletes files. It downloads Drive files into the session, uploads session files to Drive, manages sharing permissions (owner, writer, commenter, reader), and reads a file's version history.

Full operation list
  • Find — gdrive_find_file (by name, type, date, content), gdrive_find_folder
  • Create — gdrive_create_file (empty Doc/Sheet/Slide), gdrive_create_file_from_text (up to 10MB text-like content), gdrive_create_folder
  • Modify — gdrive_edit_file (overwrite binary), gdrive_copy_file, gdrive_move_file, gdrive_trash_file, gdrive_untrash_file, gdrive_delete_file
  • Download / upload — gdrive_download_file (saves to session in/), gdrive_upload_file (5MB per call)
  • Sharing — gdrive_create_permission (roles: owner, writer, commenter, reader), gdrive_list_permissions, gdrive_delete_permission
  • Versions — gdrive_list_revisions, gdrive_get_revision, gdrive_update_revision

Try asking"find the Q3 roadmap doc in my Drive, download it, and summarise the top 5 initiatives."

My GDocs (16 tools via MCP)

What it does — creates, reads, edits, and exports Google Docs with heavy markdown support. It can search your documents, read them as plain text or with full structure, and build new docs section by section with rich formatting. Edits range from simple text inserts and document-wide find-and-replace to precise programmatic changes; it also inserts and swaps images and tables, copies documents with formatting intact, and exports to PDF.

Full operation list
  • Find — gdocs_search_documents
  • Read — gdocs_get_document_plaintext (quick reads), gdocs_get_document (full structure with indices, inline objects, headers, footers)
  • Create — gdocs_create_document (empty) + gdocs_update_document_section_markdown for incremental rich formatting (more reliable than one-shot markdown); or gdocs_create_document_markdown for simple docs
  • Edit — gdocs_insert_text, gdocs_replace_all_text, gdocs_delete_content_range, gdocs_batch_update (powerful programmatic edits)
  • Images — gdocs_insert_inline_image (requires publicly accessible URL), gdocs_replace_image
  • Tables — gdocs_insert_table then gdocs_batch_update to populate cells
  • Copy / export — gdocs_copy_document (preserves images, formatting, headers/footers), gdocs_export_pdf

Try asking"create a new Google Doc titled 'Q3 Strategy', add sections for Goals, Metrics, and Risks — keep each concise."

My GSheets (17 tools via MCP)

What it does — reads, writes, formats, and manages spreadsheets. It can find and inspect workbooks, read one or many ranges, write or append values, and update-or-insert rows keyed on a column. It also cleans data (clear ranges, find-and-replace, delete rows or columns), restructures workbooks (new spreadsheets, add or remove sheets), looks up rows by exact match, and applies cell formatting and charts.

Full operation list
  • Find / inspect — gsheets_search, gsheets_get_info, gsheets_get_sheet_names
  • Read — gsheets_get_values (single range) or gsheets_batch_get (multiple). Use bounded ranges on large sheets.
  • Write — gsheets_update_values, gsheets_append_values, gsheets_upsert_rows (update-or-insert by key column, auto-adds missing columns)
  • Clean — gsheets_clear_values, gsheets_find_replace, gsheets_delete_dimension
  • Structure — gsheets_create_spreadsheet, gsheets_add_sheet, gsheets_delete_sheet
  • Find row — gsheets_lookup_row by exact cell match
  • Format — gsheets_format_cell (bold, colors, font size), gsheets_create_chart

Try asking"open the CRM sheet, find rows where the status is 'stalled', and summarise the top 5 largest stalled deals."

My GSlides (7 tools via MCP)

What it does — creates professional Google Slides presentations straight from markdown, themed with one of eight built-in looks (modern dark, corporate blue, professional gray, creative purple, warm orange, forest green, minimal beige, or the default). It can also update an existing presentation, read a deck or a single page, fetch slide thumbnails, and copy from a template deck. For decks built around local or generated images, use the built-in Presentation Create tool instead — the markdown path only accepts a few public image hosts.

Full operation list

The standout is gslides_create_slides_markdown — creates presentations from markdown with themes: modern_dark, corporate_blue, professional_gray, creative_purple, warm_orange, forest_green, minimal_beige, default.

Formatting rules for best results:

  • Start with Theme: modern_dark (or another) as the first line.
  • Title slide: # Title\nSubtitle
  • Content: ## Slide Title + bullets (8–15 words each for readability)
  • Quote: > concise quote under 120 chars
  • Two-column: left content, ||| on its own line, right content
  • Images: ![alt](url) — only Unsplash (images.unsplash.com), GitHub raw, and Google branding URLs work. For local/generated images, use the built-in Presentation Create tool instead.
  • Keep slides to 3–6 bullets each, one topic per slide, separated by \n---\n.

Also: gslides_update_presentation (markdown OR raw Slides API requests, not both), gslides_get_presentation, gslides_get_page, gslides_get_thumbnail, gslides_copy_from_template.

Try asking"create a Google Slides deck titled 'Series B Pitch' with the modern dark theme and sections for problem, solution, traction, team, and ask."

My GA4 (36 tools via MCP)

What it does — Google Analytics 4 reporting and property admin, read-only. It discovers every account, property, and data stream you can reach, then runs reports: core dimension-and-metric queries, realtime (the last 30 minutes), cross-tabs, funnels, batched bundles, and long-running jobs for big pulls. It also checks whether a metric combination is queryable and how much quota remains before a heavy pull, lists your custom dimensions, metrics, and channel groups, reads conversion and key events, and exports actual audience membership lists.

Full operation list
  • Discovery — ganalytics_list_account_summaries is the entry point (shows every account + property + data stream you can reach), ganalytics_list_accounts_v1_beta, ganalytics_list_properties_filtered, ganalytics_get_property, ganalytics_list_data_streams
  • Reporting (sync) — ganalytics_run_report for the core dimension+metric query; ganalytics_run_realtime_report for the last 30 minutes; ganalytics_run_pivot_report for cross-tabs; ganalytics_run_funnel_report for funnels; ganalytics_batch_run_reports / ganalytics_batch_run_pivot_reports to bundle up to 5 requests in one round-trip
  • Reporting (async, for long/large jobs) — create a report task, poll ganalytics_list_report_tasks until state is ACTIVE, then pull rows with ganalytics_query_report_task
  • Metadata & safety — ganalytics_get_metadata to discover valid dimension/metric apiNames before running any custom report; ganalytics_check_compatibility to verify a dimension+metric combo is queryable; ganalytics_get_property_quotas_snapshot to see remaining quota tokens before batching
  • Custom definitions — ganalytics_list_custom_dimensions, ganalytics_list_custom_metrics, ganalytics_list_calculated_metrics, ganalytics_list_channel_groups
  • Events — ganalytics_list_conversion_events, ganalytics_list_key_events
  • Audiences — ganalytics_list_audiences, ganalytics_get_audience for definitions; ganalytics_list_audience_lists, ganalytics_query_audience_list, ganalytics_list_recurring_audience_lists for actual user membership exports

Workflow rulealways call ganalytics_get_metadata first to discover valid apiNames for a property (the GA4 UI labels ≠ API names), then ganalytics_check_compatibility to avoid INVALID_ARGUMENT from mixing session-scoped and user-scoped fields, then ganalytics_run_report. Pass property as properties/{numeric_id} (12-digit).

Date ranges accept YYYY-MM-DD or relatives like today, yesterday, 7daysAgo, 30daysAgo. Realtime reports don't take dateRanges.

Heads up — every tool is read-only. Audience-list exports return actual end-user identifiers and are governed by a Safety Center toggle ("Export audience user lists").

Try asking"run a GA4 report on our main property for the last 30 days — active users, sessions, and bounce rate by country and device, sorted by sessions."

My Outlook (21 tools via MCP)

What it does — Microsoft mail + calendar on your connected account. On the mail side it searches, lists, and reads messages, handles attachments, sends email, creates and sends drafts, and replies in-thread. On the calendar side it lists your calendars and events and creates, updates, or deletes events.

Full operation list
  • Mail — outlook_search_messages (KQL search), outlook_list_messages (folder-scoped), outlook_get_message, outlook_list_attachments, outlook_send_email, outlook_create_draft, outlook_send_draft, outlook_reply_email
  • Calendar — outlook_list_calendars, outlook_list_events, outlook_get_event, outlook_create_event, outlook_update_event, outlook_delete_event

Try asking"search my Outlook inbox for emails from finance@acme.com in the last 30 days and summarise the money-related requests."

My Microsoft Teams (21 tools via MCP)

What it does — Teams chats, channels, meetings, and directory. It discovers your teams and channels, reads channel and chat messages, sends chat messages, and posts or replies in channels. It also creates and looks up online meetings and searches people, messages, and files across the workspace.

Full operation list
  • Discovery — teams_list_joined_teams, teams_list_channels, teams_get_primary_channel
  • Reading — teams_list_channel_messages, teams_list_chats, teams_get_chat, teams_list_chat_messages
  • Sending — teams_send_chat_message, teams_post_channel_message, teams_reply_channel_message
  • Meetings — teams_create_meeting, teams_list_online_meetings, teams_get_online_meeting
  • Directory — teams_list_team_members, teams_list_users, teams_search_messages, teams_search_files

Heads up — many Teams Graph endpoints don't work with personal Microsoft accounts (@outlook.com, @hotmail.com, @live.com). Teams tools work best with work/school Microsoft 365 accounts backed by Entra ID / Azure AD.

Try asking"summarise the key decisions made in the #product channel over the last two weeks."


6. Finance

Markets, filings, analysts — 4 tools.

Stocks Analyser

What it does — real-time and historical stock prices, technical indicators (RSI, MACD, moving averages), price charts.

When to reach for it — technical analysis, stock screening, identifying trends and entry/exit points, price history comparison.

Try asking"chart NVDA over the last 2 years with 50- and 200-day moving averages and RSI."

Company Financials

What it does — income statements, balance sheets, cash flow, financial ratios, and SEC filings (10-K, 10-Q).

When to reach for it — fundamental analysis, due diligence, historical financial trends, comparing financial health across companies.

Try asking"pull 5 years of income statements for MSFT and compare revenue mix trends by segment."

Analyst Views

What it does — Wall Street analyst ratings, price targets, consensus estimates, upgrades/downgrades.

When to reach for it — investment research, comparing analyst sentiment across stocks, earnings analysis.

Try asking"compare analyst consensus and 12-month price targets for NVDA, AMD, and AVGO."

PSX Market

What it does — Pakistan Stock Exchange data: real-time PSX quotes, KSE-100 and KSE-30 indices, company fundamentals, sector analytics, market breadth, dividends, and OHLCV candles across Pakistan-listed equities, ETFs, futures, and bonds.

When to reach for it — PSX portfolio tracking, KSE-100 monitoring, Pakistan sector research, Pakistani company analysis.

Try asking"show KSE-100 index performance this quarter and list the top 5 sector gainers."


7. Travel & local

Flights, maps, stays, shopping — 6 tools.

Google Flights

What it does — flight search with prices, airlines, layovers, durations, and multi-city itineraries.

When to reach for it — travel planning, comparing flight options, finding the cheapest routes, multi-leg trips.

Try asking"find me the cheapest business-class flights from London to Tokyo in the first week of June, one stop OK."

Google Maps

What it does — find places, get directions, business hours, ratings, and reviews.

When to reach for it — local business research, venue comparison, route planning, finding nearby services, competitive location analysis.

Try asking"find the top-rated coffee shops in Lisbon within walking distance of LX Factory, compare reviews."

Google Shopping

What it does — product search with prices, ratings, reviews, and retailer comparisons.

When to reach for it — price comparison, finding deals, product research before a purchase.

Try asking"compare ergonomic office chairs under $500, rank by review score and shipping speed."

Release note: Amazon Search & Amazon Product — Localised Shopping On Your Storefront.

What it does — Amazon product search on the user's local storefront, auto-routed across 21 regions: US, UK, Germany, France, Italy, Spain, Netherlands, Belgium, Poland, Sweden, Turkey, Egypt, Saudi Arabia, UAE, India, Japan, Singapore, Australia, Canada, Mexico, Brazil. Results are filtered to items deliverable to your country (amazon_domain + delivery_country set automatically from your detected location).

Filters: sort_by (featured, price_low_to_high, price_high_to_low, average_review, newest_arrivals, bestsellers), price_min / price_max (currency symbols stripped automatically), pagination via page. Pass country (ISO-2, country name, or full domain) only when the user explicitly asks for a different region ("find this on amazon.com") — leave it unset by default so the user's local Amazon wins.

When to reach for it — shopping when the user names Amazon directly, Prime + delivery checks, paginated browsing on Amazon, or when Google Shopping's retailer-mix isn't specific enough. Different from Google Shopping (aggregates many retailers) — Amazon Search goes straight to the source and exposes Prime/delivery details Google Shopping doesn't.

Try asking"find me the best-reviewed wireless mechanical keyboard under £200 on Amazon" or "search Amazon for noise-cancelling headphones, sort by price low to high, page 2".

Amazon Product

Release note: Amazon Search & Amazon Product — Localised Shopping On Your Storefront.

What it does — full product detail for one Amazon listing, by ASIN (B08N5WRWNW) or full Amazon URL. URL form auto-detects the storefront from the domain. Returns title, brand, rating, buybox (price, currency, availability), feature bullets, attributes, variants, and image gallery.

Storefront precedence: URL domain wins → explicit country override → user's detected location. Useful when the user shares an Amazon link or after amazon_search returns results and the user wants to drill into one product.

When to reach for it — a user pastes an Amazon URL, an ASIN, or asks to compare the same product across two countries' Amazons.

Try asking"what's the buybox price for https://www.amazon.de/dp/B0EXAMPLE?" or "compare ASIN B0DGYHMQXR on amazon.com vs amazon.co.uk — price, Prime, delivery".

What it does — Airbnb listings with prices, ratings, amenities, and availability.

When to reach for it — travel accommodation research, vacation planning, rental price comparisons, destination feasibility.

Try asking"find 3-bedroom Airbnbs in Barcelona for June 10–17, pool preferred, under €400/night."


8. System & orchestration

Memory, playbooks, sub-agents, plans, billing — 8 tools.

Agent Intuition

What it does — reads and writes your durable memory (facts Alfrada OS has stored about you) and domain playbooks (reusable workflows). Memory is handled by four focused tools — memory_search (search or list), memory_create, memory_update, memory_delete — and playbooks by the playbook tool (read, list, create, update, delete).

When to reach for it — you usually don't call it directly; the agent consults it automatically. You do use it when you want to edit a stored fact, delete one, or codify a new playbook from a working session.

Try asking"what do you already know about me?" or "save this workflow as a playbook I can reuse."

Choose Mode

What it does — switches a task between the single Agent and the multi-agent Swarm mid-conversation, for one turn or the rest of the session. When Alfrada OS judges a job big enough for the Swarm (or a follow-up small enough to drop back to the single Agent), it asks first with a consent banner — you approve or decline. You control this under Settings → Agent Behavior → Safety → Switching to Swarm mode (Ask, Always allow, or Never).

When to reach for it — usually agent-initiated; you can also just say "use the swarm for this."

Try asking"this is a big one — plan it in Swarm mode."

Settings Tool

What it does — lets Alfrada OS change a small, allow-listed set of app settings from chat — theme, EU data residency, auto-skip Q&A, auto-approve tool activation — and open guided connection flows: it can pop up the OAuth window to connect an integration, or open a secure form that captures a password or API key straight into your Vault so secrets never sit in the chat transcript.

When to reach for it — "switch to dark mode", "connect my Gmail", or when a workflow needs a login you have not stored yet — Alfrada OS opens the secure capture popup instead of asking you to paste a password.

Try asking"turn on dark mode and connect my Google Calendar."

Agent Smith

What it does — forges independent worker agents for deep research and multi-step tasks. Workers replicate autonomously and run in parallel. Three effort levels: Morpheus (quick lookup), Trinity (solid research), Neo (exhaustive deep dive). Available only in single-agent mode.

When to reach for it — deep research, long multi-step work, financial modelling, code/data exploration, or workstreams where one worker investigates while another produces artifacts.

Try asking"send Smith on a Neo-level deep dive into the enterprise AI market — competitive landscape, pricing models, GTM motions. Return a synthesised brief."

Manage Plan

What it does — multi-step plan management beyond a simple to-do list. Tracks progress across longer agentic runs.

When to reach for it — agent-initiated for complex multi-step workstreams. You don't typically call this directly; you see it surface as the plan the agent is executing.

Ask User

What it does — structured Q&A input-gathering. The agent pauses and asks you MCQ or free-text questions before proceeding.

When to reach for it — agent-initiated, not user-initiated. It pauses on its own when a task has an unresolvable ambiguity.

Request Tool Activation

What it does — asks for your approval to enable an optional tool for the current conversation.

When to reach for it — agent-initiated. It only fires on plans that actually have access to the tool.

Billing

What it does — checks your subscription status, browses plans, buys credits, triggers upgrades or cancellations.

When to reach for it — when you want to know how much of your budget you've used, buy a credit pack, or change plan without leaving the chat.

Try asking"what plan am I on and how much of my token budget did I use this month?"

Behind the scenes. Two catalog tools are internal and never something you ask for: Use Activated Tool (the dispatcher that runs a tool you approved mid-conversation before it formally joins the session's toolset — you only ever see its name in the activity feed) and Argus Schedule Self (Argus-only — lets the WhatsApp agent's heartbeat sleep until scheduled work is ready; see Argus On WhatsApp).


How To Ask For The Right Tooling

Don't name internal tools — describe the job:

  • "Search the web and compare recent vendor positioning."
  • "Run a spreadsheet-style analysis on this CSV."
  • "Generate a polished board deck from this report."
  • "Pull context from my calendar and draft a reply."
  • "Search my previous conversations and find the file we used last week."
  • "Create a recurring monitor and send me a finished update each week."
  • "Use my connected tools to draft the document, update the sheet, and send the follow-up."

Alfrada OS picks the right capabilities. This page exists so you know what's possible — not so you have to memorise it.

For copy-ready prompts per tool, see the Tool Prompt Library. For complete workflows chaining multiple tools, see the Playbook Library.

Built for Alfrada OS.