EveryDev.ai
Subscribe
Home
Tools

4,072+ AI tools

  • New
  • Trending
  • Featured
  • Compare
  • Arena
Categories
  • Agents2782
  • Coding1973
  • Infrastructure825
  • Projects603
  • Marketing598
  • Research520
  • Analytics468
  • Design462
  • MCP419
  • Testing346
  • Security323
  • Data305
  • Integration224
  • Prompts220
  • Communication210
  • Extensions196
  • Learning179
  • Voice175
  • Commerce160
  • DevOps135
  • Web95
  • Finance31
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Tools
    3. jev-effort
    jev-effort icon

    jev-effort

    AI Coding Assistants

    A proxy tool that uses Jev to dynamically select Claude Code's reasoning effort per step without breaking the prompt cache, reducing costs at max effort by up to 55%.

    Visit Website

    At a Glance

    Pricing
    Open Source

    Free and open-source under the MIT License. Run via npx with no cost beyond API usage.

    Engagement

    Available On

    API
    CLI

    Resources

    WebsiteDocsGitHubllms.txt

    Topics

    AI Coding AssistantsCompute OptimizationAgent Frameworks

    Alternatives

    skills-for-humanityClaudeGateDiagram Design
    Developer
    ifoster01ifoster01 is an independent developer on GitHub who built je…

    Listed Sep 2026

    About jev-effort

    jev-effort is an open-source research project by GitHub user ifoster01 that runs Claude Code behind a local proxy to dynamically adjust reasoning effort on every agent step using Jev, TypeSafe's small decision model. The project was created to answer a concrete question: does letting Jev choose Claude Code's reasoning effort step by step actually save money, and by how much?

    What It Is

    jev-effort is a CLI proxy tool written in JavaScript that intercepts Claude Code's API requests via ANTHROPIC_BASE_URL. Before each model request, it sends Jev a trimmed view of the conversation and asks which effort level the next step needs and for how many steps to hold it. It then applies the answer by appending an effort-only system message — using Anthropic's per-message effort beta — without resetting the prompt cache. The tool is unofficial and not affiliated with Anthropic or TypeSafe.

    How the Proxy Works

    The proxy sits between Claude Code and the Anthropic API. On each step it:

    • Sends Jev a condensed view of the conversation (prompts, Claude's visible replies, the last six tool calls)
    • Receives Jev's effort recommendation and duration
    • Appends an effort-only system message as the last message in the request (required for the beta per-turn control to take effect)
    • Re-inserts earlier effort messages at their original positions so the cached prefix remains intact

    Jev may lower effort but never exceed the session's own ceiling setting. The tool also supports a shadow mode (--jev-shadow) that logs Jev's choices without changing anything, enabling safe observation on real work.

    Cache Preservation: The Key Technical Finding

    A central concern was whether changing effort mid-session would invalidate Claude Code's prompt cache. The project confirmed it does not have to: Anthropic's per-message effort beta allows an effort-only system message to change the level while leaving the cached prefix intact. In a 470-step, 92-minute real session, the tool recorded 0 unexpected cache misses in 411 checked steps and a 99.1% cache hit rate, even with effort changing on every step.

    Benchmark Results

    The project ran 48 headless Claude Code sessions across six small coding tasks, comparing fixed effort against Jev-chosen effort under the same ceiling:

    • At high effort: Cost fell about 1.1%, thinking tokens dropped 46%, all hidden tests passed. Opus 5.5 at high already thinks little on routine steps, leaving little to cut.
    • At max effort: Cost fell 55% (21% excluding one outlier task), thinking tokens dropped 98.5%, wall time fell 58%, all hidden tests passed. Jev moved routine steps to low or medium, where max would otherwise think heavily even on simple work.
    • Real session at max (shadow mode): Estimated saving of 4–9%, because cached context reads accounted for 82% of spend and effort does not affect those costs.

    Tradeoffs and Limitations

    The README is explicit about where the approach falls short:

    • At high effort, the saving is marginal (~1%) because thinking is already a small share of the bill
    • In long sessions, context re-reading dominates cost regardless of effort level
    • Jev adds ~601 ms median latency per decision, though at max the reduced thinking more than compensates
    • The benchmark tasks are small and well-specified; quality impact on complex real work is untested
    • Using any custom ANTHROPIC_BASE_URL disables Claude Code's MCP tool search by default (workaround: ENABLE_TOOL_SEARCH=true)
    • One real session at one effort level is the only real-world measurement

    Current Status

    The repository was created and last updated in late September 2026. The code is described as working and results as reproducible. It is an MIT-licensed research project available via npx jev-effort, with commands for shadow mode, spend analysis, and controlled benchmarking. Related projects in the same space include jev-opus, jev-model-router, and effort-router, none of which (as of the README's writing) reported total cost against a fixed-effort baseline.

    jev-effort - 1

    Community Discussions

    Be the first to start a conversation about jev-effort

    Share your experience with jev-effort, ask questions, or help others learn from your insights.

    Pricing

    OPEN SOURCE

    Open Source

    Free and open-source under the MIT License. Run via npx with no cost beyond API usage.

    • Per-step effort selection via Jev
    • Shadow mode logging
    • Spend analysis (jev-effort stats)
    • Controlled benchmark runner
    • Cache-safe effort injection

    Capabilities

    Key Features

    • Per-step reasoning effort selection via Jev decision model
    • Local proxy via ANTHROPIC_BASE_URL without breaking prompt cache
    • Shadow mode to log Jev's choices without changing behavior
    • Spend analysis with jev-effort stats command
    • Controlled benchmark runner (jev-effort bench)
    • Cache-safe effort injection using Anthropic per-message effort beta
    • Effort ceiling enforcement (Jev never exceeds session setting)
    • Support for low, medium, high, xhigh, and max effort levels
    • MCP tool search compatibility via ENABLE_TOOL_SEARCH=true
    • MIT licensed and reproducible results

    Integrations

    Claude Code
    Anthropic API
    Jev (TypeSafe decision model)
    OpenRouter (Jev model access)
    Claude Opus 5.5
    Claude Fable 5.1
    Claude Mythos 5.1
    Claude Opus 5
    API Available
    View Docs

    Ratings & Reviews

    No ratings yet

    Be the first to rate jev-effort and help others make informed decisions.

    Developer

    ifoster01

    ifoster01 is an independent developer on GitHub who built jev-effort, a research proxy tool for dynamically adjusting Claude Code's reasoning effort per step using the Jev decision model. The project focuses on cost optimization for AI coding agent sessions and includes controlled benchmarks and real-session measurements to validate its findings.

    Read more about ifoster01
    WebsiteGitHub
    1 tool in directory

    Similar Tools

    skills-for-humanity icon

    skills-for-humanity

    171 structured reasoning methodologies from history's most rigorous thinkers, packaged as Claude Code skills across 27 categories.

    ClaudeGate icon

    ClaudeGate

    A high-performance local API gateway that bridges Claude Code CLI and Anthropic SDKs to any OpenAI-compatible LLM provider, with zero-crash streaming, multi-provider failover, and PII redaction.

    Diagram Design icon

    Diagram Design

    A Claude Code skill that generates 27 types of editorial-quality, brand-matched diagrams as self-contained HTML/SVG files — no Figma, no generic templates.

    Browse all tools

    Related Topics

    AI Coding Assistants

    AI tools that help write, edit, and understand code with intelligent suggestions.

    894 tools

    Compute Optimization

    Tools for optimizing computational resources and performance.

    40 tools

    Agent Frameworks

    Tools and platforms for building and deploying custom AI agents.

    754 tools
    Browse all topics
    Back to all toolsSuggest an edit
    ratings
    discussions