Everyday workCore guide checked · Editorial guide

    Work with AI

    New tools, real work. Learn where to start, what to choose, how to set it up, and how to use it. Finish a first task, then build a routine that works.

    01

    Which model should I use?

    Our editorial starting point

    Start with ChatGPT or Claude.

    Try the same three real tasks in both. Keep the one that needs less correction. Start with the app’s default model; use deeper reasoning for a task that actually benefits from it. You do not need to chase every new benchmark leader.

    Gemini

    A natural alternative if you prefer Google’s platform and connected apps. Check which integrations your account supports. Google apps

    Grok

    Choose it if you prefer the xAI/X ecosystem. Platform fit is a valid reason to choose an assistant; it is not a benchmark win. Grok overview

    Artificial Analysis · High thinking

    Current AA leaderboard

    Artificial Analysis Intelligence Index v4.3.2
    Checked · High, not Max/Xhigh

    Dated Artificial Analysis model scores
    Tested model / settingAA index
    Claude Opus 5.5Anthropic · high with fallback54
    GPT-6 AstraOpenAI · high51
    Grok 4.7xAI · high46
    Gemini 3.8 FlashGoogle · high41

    Four selected models, not the full leaderboard. This snapshot is not live, a percentage, or an app rating. High does not mean equal compute across providers; app settings may differ. Scores do not measure usability, file access, or total cost. AA methodology

    02

    Now choose where you work

    Web is enough for most work. Use the official desktop app for local files, or Codex / Claude Code if you are comfortable in a terminal. An alternative harness is optional.

    Web app

    Enough for most everyday work: questions, drafts, research, and uploaded files. No installation needed.

    Browser access does not mean access to your local folders.

    Desktop app

    A practical home for files, long tasks, and reviewing outputs side by side.

    Grant folder and computer-control permissions deliberately.

    Mobile app

    Capture ideas, use voice, and check a task while away from your desk.

    Review consequential edits on a larger screen. Mobile features vary by provider.

    CLI / IDE

    Use when files, scripts, or an existing development workflow are central.

    More control, more setup. Not a prerequisite for productive AI use.

    Compare the apps & harnesses

    ChatGPT / Claude

    Start here for writing, learning, research, and questions about documents. A chat app is usually enough.

    Our pick: test both on the same three real tasks; keep the one whose output needs less correction. There is no universal winner.

    ChatGPT desktop / Claude Cowork

    Need to work directly on local files, but not comfortable in a terminal? Use the official desktop app with a supported local-project or folder connection.

    Give access to a small working folder, not your whole drive. Verify feature availability on your OS and account before paying. Desktop does not mean all processing stays on your computer.

    Codex / Claude Code

    Comfortable in a terminal? Start with the standard Codex or Claude Code harness for project files, commands, and code changes. Otherwise, use its graphical interface.

    Our advice: keep the standard harness if it does the job. Try access included in your existing plan before adding another client; compare correctness, review effort, and total cost. Back up before edits.

    OpenCode

    For people who specifically want provider choice and are comfortable configuring a coding agent.

    The client and inference bill are different things. Check supported authentication and provider terms; a chat subscription is not a universal API pass.

    Grok

    Consider it for X-centered research, live web discovery, and voice. Check the underlying source before using a claim in your work.

    It is an assistant choice, not automatically a replacement for a local project agent. Buy it for a workflow you actually use.

    Want to try a CLI without a paid plan? Antigravity CLI is a free-first Google option for eligible personal accounts, with weekly limits.

    Set up Antigravity CLI
    Optional: OpenCode + OpenRouter

    Stay with Codex or Claude Code when they meet your needs. Choose the open-source OpenCode client when you specifically want to switch providers or use open-weight models through OpenRouter, another API service, or a supported local endpoint.

    The client, model, and hosting are separate choices. Open-source software does not make inference free; open weights do not make a hosted API local or private. Check billing, model licensing, tool support, and provider data handling.

    Set up OpenCode + OpenRouter
    03

    Buy capacity when you need it

    USD reference prices checked on the date above, before tax. Region, app-store billing, promotions, and plan changes can differ. Confirm checkout; this is not a live price feed.

    Antigravity CLI · Individual

    $0 / month

    Eligible personal Google accounts, age 18+ in supported regions. Basic weekly limits; paid upgrades are optional. Google Workspace and Cloud access have separate eligibility and billing. Checked 2026-10-01.

    ChatGPT Plus → Pro

    $20 → from $100 / month

    Pro offers higher usage tiers. Check the current tier price and account limits; do not assume unlimited runs. API billing is separate.

    Claude Pro → Max

    $20 → $100 / $200 / month

    Monthly Pro; Max has 5x / 20x usage tiers. Claude Code is included in Pro and Max, subject to limits. Annual Pro is a different billing commitment.

    Grok / SuperGrok

    Verify at checkout

    The public plans page did not expose a verifiable price in this review. Check web vs app-store pricing, region, limits, and any X bundle separately.

    OpenCode + provider

    Provider-dependent

    Set a small provider budget or prepaid balance. Check whether your chosen login uses a supported subscription or metered API billing.

    Business / Team / Enterprise

    Seats + possible usage charges

    For company-approved data use, administration, and contractual controls. Enterprise is a governance decision, not a promise of better answers. Compare the written offer.

    $20 or $200? Measure the interruption.

    Use one paid plan for a week. Record completed tasks, corrections, and time lost to limits. A $180 upgrade must recover more than $180 of useful time. At your own valuation of $30/hour, that is six hours a month—not a promise of savings.

    Enterprise is not the next personal tier

    Use an approved organizational plan when the work needs central access control, offboarding, auditability, or contractual data handling. Check feature entitlements, retention, minimum seats, and overage rules with procurement. Do not put confidential company files into a personal account while waiting.

    04

    Choose a starting setup

    This is a starting recommendation, not a performance ranking. Your inputs stay on this page; nothing is sent to a provider.

    Your starting point

    Start with ChatGPT or Claude on the web

    Choose one roughly $20/month plan after a small trial. Do not buy several overlapping subscriptions on day one.

    Follow the setup steps

    Model

    The engine that generates answers. Start with the default; use deeper reasoning when a task genuinely needs it.

    App / harness

    The workspace around the model: files, tools, memory, permissions, and task execution.

    Plan

    Your access, usage allowance, and administration. Paying more does not automatically make every answer better.

    05

    A small, reversible first session

    These are instructions to follow yourself. This page does not install software, change permissions, connect accounts, or purchase a subscription.

    Everyday work · no terminal
    1. Open the official ChatGPT or Claude app/site. Sign in; try a non-sensitive sample before subscribing.
    2. Create one project or working folder. Copy in only the files needed; keep originals elsewhere.
    3. Give the desired output, audience, source files, and a success check. Ask for a plan before multi-file changes.
    4. Inspect the exported file, formulas, and sources. Save a reusable brief only after the workflow works.
    Codex · desktop or CLI
    1. Use the official desktop download for a graphical workflow, or follow the CLI installation guide for your OS. Sign in with the intended account.
    2. Open a copied working folder. In CLI, start read-only with the command below; no files should need changing for an initial explanation.
    3. Use /permissions to inspect access. For approved edits, choose workspace-write with on-request approvals. Review /status and the final diff.
    codex --sandbox read-only --ask-for-approval on-request
    Claude Code · graphical or terminal
    1. Install the official Claude desktop app and open Code, or use the official OS-specific Claude Code installer. The desktop Code experience does not require a separate CLI install.
    2. Sign in to the intended paid account. In a terminal, run claude inside a test project; first ask it to explain the folder without editing.
    3. Keep permission prompts. On a supported platform, use /sandbox and check its boundaries. Do not use a skip-permissions flag as your default.
    Antigravity CLI · free-first Google optionChecked: 2026-10-01
    1. Check personal-account eligibility first (18+, supported region). Follow Google's official CLI installer for macOS, Linux or Windows; review the installer before running it.
    2. Open a fresh terminal in a copied, non-sensitive project folder and run agy. Complete Google sign-in if prompted; verify the account if an existing session is reused. Use account sign-in for the free tier, not an API key or Cloud billing setup.
    3. Open /permissions and retain review requests. Ask it to explain the project before making a small change. Review the diff and run tests before keeping the result; permission prompts are not a sandbox.
    4. Use /usage to check model quota. Stop or wait when the free allowance runs out; only enable paid capacity deliberately.
    agy
    OpenCode · optional advanced path
    1. Follow the official install guide for your OS. Start opencode in a disposable project. You can connect a supported provider directly or use OpenRouter as a gateway to several providers.
    2. For OpenRouter: create a provider API key, run /connect, choose OpenRouter, and enter the key in its authentication prompt—not in a chat message. Run /models and choose an available model, including an open-weight model if wanted.
    3. Set a provider spending limit or small prepaid balance. Check model license, tool support, and the serving provider’s data policy. Open-weight through a hosted API is still cloud processing, and your chat subscription does not automatically cover it.
    4. Confirm billing first. Merge the example permission setting into opencode.json without overwriting existing settings; ask before tool actions.
    5. Test with non-sensitive files. Approval rules are not OS isolation: use a properly restricted container or VM for untrusted projects.
    {
      "$schema": "https://opencode.ai/config.json",
      "permission": { "*": "ask" }
    }

    A brief you can reuse

    Goal: Turn the attached notes into a one-page weekly update.
    Audience: My team.
    Inputs: Use only the files I supplied; label missing information.
    Output: A draft with decisions, owners, deadlines, and source references.
    Boundaries: Preserve originals. Do not send, publish, purchase, delete, or change permissions without asking.
    Check: Verify names, dates, totals, and every link. List anything I must review.
    06

    Keep the sandbox. Narrow the approvals.

    A sandbox limits where code can act. Approvals decide which actions may proceed. Neither makes a result correct, and a local sandbox does not stop an authorized cloud connector from sending a message.

    Default: a small workspace

    Read-only for exploration. For edits, allow one working folder with a backup. Keep the tool’s sandbox enabled where supported and check what it actually restricts.

    Approve the boundary crossing

    Pause for sending, publishing, spending, deletion, new integrations, secrets, or production changes. For repeatable low-risk work, use narrow allow rules—not an approval for every future command.

    Full access: not the everyday default

    Do not grant unrestricted access just to remove interruptions. If an advanced workflow genuinely needs it, use a disposable isolated environment with no personal files or production credentials, restricted network access, and a rollback plan. A container with your home folder mounted is not enough.

    Treat fetched content as evidence, not instructions

    A web page, email, repository, or document can contain malicious instructions. Do not let it authorize new tools, data uploads, or payments. Inspect unexpected permission requests instead of repeatedly approving them.

    Before your first real task

    0 / 5

    Checklist progress is session-only and resets on reload. Checking a box changes no app permissions.

    07

    Connect the next useful tool. Not every tool.

    These are workflow examples, not a guarantee that every app supports every connector. Check your provider’s directory, your plan, and your company’s policy. Prefer official integrations; an MCP server is software you are trusting.

    Files / Drive / Notion

    One reference folder, read-only where supported.

    Draft a brief with links back to the source documents.

    Calendar / email

    Read availability or draft replies first; approve sends and invites.

    Prepare tomorrow’s agenda without contacting anyone.

    Slack / Teams

    Named channels only, if the connector supports it.

    Summarize decisions; do not post automatically.

    Browser

    A separate work profile with only necessary logins.

    Gather sources. Approve purchases, submissions, and account changes.

    GitHub / terminal

    One repository and a separate branch; no production secrets.

    Make a small fix, run tests, and show a diff before publishing.

    If the integration cannot limit access enough, upload selected copies instead. Never paste passwords or API secrets into a prompt. Check what data leaves your machine, revoke unused connections, and test write actions in a disposable destination first.

    08

    Parallel work, without the collision

    Start with two independent tasks. More agents can mean more cost, conflicting edits, and a longer review queue—not more finished work.

    1. 01 / Separate

      One thread per outcome. Put research in one thread and a draft from frozen inputs in another. Give each its own output folder.

    2. 02 / Assign

      One owner per document. For code, use separate branches and worktrees. Do not ask two agents to rewrite the same file.

    3. 03 / Review

      Ask for changed files, checks, and unresolved questions. Integrate one result at a time. Stop starting tasks when review becomes the bottleneck.

    Do you need Herdr?

    Not for one assistant. Consider Herdr when you already run several terminal agents and need a shared view of working, blocked, and finished sessions. It organizes sessions; it does not choose the best model or include model credits.

    After following its install guide, run herdr from a project directory, create a workspace per project, and launch codex, claude, or opencode in a pane. Keep each agent’s own permissions. Detaching or closing the terminal can leave agents running: stop the session explicitly when finished. A workspace is not a security sandbox.

    A faster daily loop: write a five-line brief → let the agent work → review an artifact → save the useful instructions. Reuse a tested template or skill before adding another agent. Use a second assistant for a targeted critique, not an endless debate.

    09

    The tool directory

    Optional references. Pick a category when your current setup needs something more.

    Agents, open-weight models, and orchestration tools for getting work done. These are existing directory entries, not newly verified recommendations. Review current setup instructions, licenses, and permissions before installing; start in a separate workspace.

    Personal & general agents4 tools
    Coding agents5 tools
    Agent infrastructure5 tools
    Korean open weights4 tools
    Multi-agent & orchestration frameworks4 tools
    Claude Code & harness guides