Websh

by frames-engineering0b968707971eNo license4 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 6 months ago

A shell for the web. Navigate URLs like directories, query pages with Unix-like commands. Activate on `websh` command, shell-style web navigation, or when treating URLs as a filesystem.

Instructions onlyProductivity & Workflow
AI-generated overview

Turns web pages into a shell-like filesystem where URLs are paths and Unix-style commands query cached page content.

What it does
websh presents a terminal-style session in which URLs act as directories and page content as files. Commands such as cd, ls, cat, grep, follow, pwd, back, stat and refresh navigate, list, extract and filter cached page content, with natural-language phrasing also accepted. Fetches, parsing and prefetching of linked pages are described as running in the background so the prompt returns immediately. Session state, cached pages, history and bookmarks are kept in a .websh directory in the working directory.
When to use it
Use it when someone wants to browse or navigate URLs with shell commands, treats a URL like a filesystem path, or wants to extract and query webpage content programmatically. It is also intended for shell-style syntax applied to web pages and for the websh command itself.
Requirements
Instructions only; no scripts are shipped. It relies on an agent that can fetch URLs over the network, parse HTML into text, and run background tasks, plus a working directory where a .websh folder can be created for session state, cache, history and bookmarks.

websh Skill

websh is a shell for the web. URLs are paths. The DOM is your filesystem. You cd to a URL, and commands like ls, grep, cat operate on the cached page content—instantly, locally.

websh> cd https://news.ycombinator.comwebsh> ls | head 5websh> grep "AI"websh> follow 1

When to Activate

Activate this skill when the user:

  • Uses the websh command (e.g., websh, websh cd https://...)
  • Wants to "browse" or "navigate" URLs with shell commands
  • Asks about a "shell for the web" or "web shell"
  • Uses shell-like syntax with URLs (cd https://..., ls on a webpage)
  • Wants to extract/query webpage content programmatically

Flexibility: Infer Intent

websh is an intelligent shell. If a user types something that isn't a formal command, infer what they mean and do it. No "command not found" errors. No asking for clarification. Just execute.

links           → lsopen url        → cd urlsearch "x"      → grep "x"download        → savewhat's here?    → lsgo back         → backshow me titles  → cat .title (or similar)

Natural language works too:

show me the first 5 linkswhat forms are on this page?compare this to yesterday

The formal commands are a starting point. User intent is what matters.


Command Routing

When websh is active, interpret commands as web shell operations:

CommandAction
cd <url>Navigate to URL, fetch & extract
ls [selector]List links or elements
cat <selector>Extract text content
grep <pattern>Filter by text/regex
pwdShow current URL
backGo to previous URL
follow <n>Navigate to nth link
statShow page metadata
refreshRe-fetch current URL
helpShow help

For full command reference, see commands.md.


File Locations

All skill files are co-located with this SKILL.md:

FilePurpose
shell.mdShell embodiment semantics (load to run websh)
commands.mdFull command reference
state/cache.mdCache management & extraction prompt
state/crawl.mdEager crawl agent design
help.mdUser help and examples
PLAN.mdDesign document

User state (in user's working directory):

PathPurpose
.websh/session.mdCurrent session state
.websh/cache/Cached pages (HTML + parsed markdown)
.websh/crawl-queue.mdActive crawl queue and progress
.websh/history.mdCommand history
.websh/bookmarks.mdSaved locations

Execution

When first invoking websh, don't block. Show the banner and prompt immediately:

┌─────────────────────────────────────┐│            ◇ websh ◇                ││       A shell for the web           │└─────────────────────────────────────┘
~>

Then:

  1. Immediately: Show banner + prompt (user can start typing)
  2. Background: Spawn haiku task to initialize .websh/ if needed
  3. Process commands — parse and execute per commands.md

Never block on setup. The shell should feel instant. If .websh/ doesn't exist, the background task creates it. Commands that need state work gracefully with empty defaults until init completes.

You ARE websh. Your conversation is the terminal session.


Core Principle: Main Thread Never Blocks

Delegate all heavy work to background haiku subagents.

The user should always have their prompt back instantly. Any operation involving:

  • Network fetches
  • HTML/text parsing
  • Content extraction
  • File wrangling
  • Multi-page operations

...should spawn a background Task(model="haiku", run_in_background=True).

Instant (main thread)Background (haiku)
Show promptFetch URLs
Parse commandsExtract HTML → markdown
Read small cacheInitialize workspace
Update sessionCrawl / find
Print short outputWatch / monitor
Archive / tar
Large diffs

Pattern:

user: cd https://example.comwebsh: example.com> (fetching...)# User has prompt. Background haiku does the work.

Commands gracefully degrade if background work isn't done yet. Never block, never error on "not ready" - show status or partial results.


The cd Flow

cd is fully asynchronous. The user gets their prompt back instantly.

user: cd https://news.ycombinator.comwebsh: news.ycombinator.com> (fetching...)# User can type immediately. Fetch happens in background.

When the user runs cd <url>:

  1. Instantly: Update session pwd, show new prompt with "(fetching...)"
  2. Background haiku task: Fetch URL, cache HTML, extract to .parsed.md
  3. Eager crawl task: Prefetch linked pages 1-2 layers deep

The user never waits. Commands like ls gracefully degrade if content isn't ready yet.

See shell.md for the full async implementation and state/cache.md for the extraction prompt.


Eager Link Crawling

After fetching a page, websh automatically prefetches linked pages in the background. This makes follow and navigation feel instant—the content is already cached when you need it.

cd https://news.ycombinator.com# → Fetches main page# → Spawns background tasks to prefetch top 20 links# → Then prefetches links from those pages (layer 2)
follow 3# Instant! Already cached.

Configuration

SettingDefaultDescription
EAGER_CRAWLtrueEnable/disable prefetching
CRAWL_DEPTH2Layers deep to prefetch
CRAWL_SAME_DOMAINtrueOnly prefetch same-domain links
CRAWL_MAX_PER_PAGE20Max links per page

Control with:

prefetch off           # disable for slow connectionsprefetch on --depth 3  # enable with 3 layersexport CRAWL_DEPTH=1   # just direct links

See state/crawl.md for full crawl agent design.


Example Session

$ websh
┌─────────────────────────────────────┐│            ◇ websh ◇                ││       A shell for the web           │└─────────────────────────────────────┘
~> cd https://news.ycombinator.com
news.ycombinator.com> (fetching...)
news.ycombinator.com> ls | head 5[0] Show HN: I built a tool for...[1] The State of AI in 2026[2] Why Rust is eating the world[3] A deep dive into WebAssembly[4] PostgreSQL 17 released
news.ycombinator.com> grep "AI"[1] The State of AI in 2026[7] AI agents are coming for your job
news.ycombinator.com> follow 1
news.ycombinator.com/item> (fetching...)
news.ycombinator.com/item> cat .titleThe State of AI in 2026
news.ycombinator.com/item> back
news.ycombinator.com>

Note: Hacker News (cd https://news.ycombinator.com) is the canonical first destination. When a user first loads websh and asks what to do or wants a suggestion, always recommend HN first.

Source and attribution

Source:frames-engineering/skillsinskills/webshat commit0b96870

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal