PANE

Claude Hits Token Limits Mid-Task With No Warning

Heavy Claude users struggle with usage limits, unexpected interruptions, and session degradation (context rot). They are actively seeking workarounds and custom solutions to optimize token usage and maintain session stability for coding tasks. This indicates a need for better control and predictability within the Claude environment.

aiproductivitycreator-economycodingchatbots
FIT
0%
SIGNAL
100%
SOURCES60
FRESHEST POST11H AGO
TRACKED SINCE106D AGO

SOURCES (60)

what's the advantage of using Hermes to do this, over just directly doing it with Claude/ChatGPT?

r/LocalLLaMA11h ago

This can't be true forever. Unlimited frontier model use, reviewing 100,000 large documents at a time, with agentic use? Otherwise, base Claude's redlining capability is improving so I remain skeptical what the incremental value is really worth.

r/legaltech11h ago

> Same goes for memory management: you can do it manually, and it might work out most of the time, but if the programming language is helping you with it, the likelihood of memory safety issues decreases drastically. And the problem with avoiding these issues only by "meeting certain standards of rigour" is that you might only find out that you have failed to meet them years later, when you learn that your software has a vulnerability.Zig does help you. Array slices, explicit nullability of poin

HN16h ago
Source preview · community.openai.com

I created a status companion for Codex and Claude Code

community.openai.com1d ago

>It's now become so bad, I had to literally vibecode a separate linter for this.You see, there's your problem right there. You're vibe coding, which by definition literally means you're unwilling to look at the generated code. That's not what successful ai assisted software developers are doing. YOU HAVE TO READ THE CODE. Refusing to do that means you're not a serious programmer, you're outsourcing your thought and design and implementation, trying to get something for nothing by taking the easy

HN2d ago

Did claude code start randomly creating artifacts for you too?Previously I was asking it to create a local html file with a summary (when needed, this isn't in any of my CLAUDE.md files).Now it creats them by itself when it sees fit, and they're stored online on claude.ai, can this be disabled?

HN2d ago

Solo side project. Public alpha. Looking for feedback from people who actually run Claude Code / Cursor / Codex / Antigravity. Problem I hit: I kept trying to "make the agent safe" with deny lists and block modes. Hooks run as me anyway. What I really needed was a black box every tool call, in order, on disk I own. What I built: Agentmetry - Hooks into Claude Code (also Cursor / Codex / Antigravity) - Writes hash-chained JSONL locally (MITRE tags, arg hashes) - Sequence detections (e.g

r/SideProject2d ago

TL;DR - Disabilities took away my ability to both manage and explore new services for my homelab. Claude code gave me that back I wouldn't be able to manage my homelab without claude code. I started my homelab journey over 10 years ago with a single tower running plex. It grew over the years to a docker compose stack of ~50 containers that I was running in my free time. Then long-covid hit me hard. It caused other disabilities to get much worse (EDS, POTS, MCAS), and I was no longer able to

r/selfhosted2d ago

lol, you just don’t know that people are using Claude. With the CLI, you basically have superpowers and can get anything done you need to do. It’s cute to use Sidekick to run a report, you can use Claude to build an entire reporting system that alerts you to changes, highlights opportunities and builds plans to take advantage of those opportunities.

r/shopify3d ago

Seeing this a lot where folks want to move faster but throw all best practices out. I recommend adding a universal secure connector. Claude <------> secure universal connector <------> talks to all work apps you want to allow. From there, your Claude can control policies to allow users to read data from their own Sharepoint account and other apps. It can even add security policies to prevent stupid stuff like global deletions or emailing a competitor's company as well as give you

r/sysadmin3d ago

Token management is a big deal when you have client projects to deliver. I need to have a pulse on what my current burn looks like, especially because I am spawning a bunch of subagents.Agents need to know in real-time what the current utilization is. I went through 10 iterations of this tool - and likely going to do a number of more iterations. The net net is that I needed something to show me if I am on pace to use the tokens alloted to my subscription.Using up the session (maxxing out my sess

HN3d ago
Source preview · reddit.com

Wait… making money is a thing?

reddit.com4d ago

I was iterating on a few prompt workflows and noticed something odd where small prompt tweaks are causing bigger cost shifts than expected since token counts aren’t changing that much on paper and outputs look similar length wise and behavior is also mostly the same but still cost per request seems to be going up. From what i've seen the only real differences are slight wording changes and some added structure for better outputs so no major model switches or obvious jumps in usage and at thi

r/PromptEngineering4d ago

I really hate that they added mouse integration with Claude Code.Nothing like it prompting you for an answer and you click on to the terminal accidentally, resulting in you choosing an answer.

HN4d ago
Source preview · reddit.com

since nobody paid more for more efficiency .... no

reddit.com5d ago

Hi HN, i am Dimitris,I have been using Claude Code and Codex agents, for some time now from the beggining i had been using them from inside my terminal mainly for coding.For the past 3 years i kept building so i have a homelab and a business server, which already has a lot of vms, so a lot of things to manage.So i decided to start building my own ideal version of using claude code, and codex flexibly and connect them easily with all my vms and infra but at the same time keep using them for codin

HN5d ago

Conversation transcribing + prd building directly with Claude code and then shipping it directly to Jira via MCP is the biggest unlock for me as well

r/ProductManagement5d ago
Source preview · reddit.com

submitted by /u/CommunityTechnical99 [link] [comments]

reddit.com5d ago
Source preview · reddit.com

Thanks, this was a lifesaver :)

reddit.com5d ago

I kept finding myself in situations where Claude Code would churn on easy tasks. There’s a lot of advice online about what to do, but very little on why it works. I wanted to cut through the noise, so I went looking for the connection between the best practices and how these models actually work at a fundamental level. Context turned out to be the thing to anchor on. Once you have the intuition for what a session is actually doing with everything you’ve let into the window, the habits stop being

r/datascience6d ago

Great thoughts. Yeah, we require turning Claude memory off to prevent cross-matter contamination. And we obviously use data minimization or anonymization practices to try to reduce the amount of confidential data we put in, because why not. Routine deletion is good too. Glad to see someone else reaching the same conclusions! Totally agree re the Heppner overreaction.

r/legaltech6d ago
Source preview · reddit.com

How did you do the hallucination check?

reddit.com6d ago

i am in similar boat, my workplace is very locked down too. no admin rights on my machine so i cant install anything myself but what i do is use the tools at home for personal projects, sure is not "actual workflow" like they ask but you still learn how to prompt properly and what the tools are good at. the skills transfer more than you think also some companies just write that requirement but they dont really enforce it strictly. if you show you understand the concepts and can talk ab

r/cscareerquestions7d ago

Yeah, that's exactly the case it's built for. Instead of staying silent until you're out, it pings you on the way there, at 25, 50, 75, 90 and 95 percent of the 5 hour window. So you get a heads up at 75 or 90 percent while you can still decide to wrap up or push the big job to later, rather than finding out by getting throttled. The thresholds are the part I'm still tuning though, which is why I asked. Would something like a single nudge at 90 percent be enough for you, or would

r/SideProject7d ago
Source preview · towardsdatascience.com
https://towardsdatascience.com/governed-context-managing-context-rot-in-claude-code/
towardsdatascience.com7d ago

Yeah, Claude found it for me while I was debugging my vLLM config. This needs pinning somewhere.

r/LocalLLaMA7d ago

To be clear, it’s still a manual workflow. One session can’t directly command another. What I do instead is use a dedicated “controller session” that reviews the relevant transcripts alongside the current specs, reconstructs the intent and progress of each session, and then tells me what should land first versus what can safely run in parallel. If a session was closed, it can even identify which session ID I should resume first. I’ve found that keeping the active session count around three or fo

r/SideProject7d ago

This is how I worked with LLMs originally, and I much preferred it. This gave me a much better understanding of the code that I was adding. But, there's no way to keep up with my team like this anymore. It's just too slow when everyone else is working directly in Claude Code.

HN8d ago

for *nix systems, add these to your shell's init. (e.g. ~/.bashrc) export MODEL_NAME="qwen3.6" export ANTHROPIC_AUTH_TOKEN=ollama export ANTHROPIC_API_KEY="" export ANTHROPIC_BASE_URL=http://localhost:8000 export ANTHROPIC_DEFAULT_HAIKU_MODEL=$MODEL_NAME export ANTHROPIC_DEFAULT_SONNET_MODEL=$MODEL_NAME export ANTHROPIC_DEFAULT_OPUS_MODEL=$MODEL_NAME export CLAUDE_CODE_SUBAGENT_MODEL=$MODEL_NAME export CLAUDE_CODE_AUTO_COMPACT_WINDOW=262144 export CLAUDE_AUTOCOMPACT_PCT_O

r/LocalLLaMA8d ago

How do I get this to work with Claude? Maybe Cloudflare is the issue.

r/selfhosted8d ago

Sonn is a one-line terminal install that adds a persistent reasoning layer to Claude Code. It automatically captures project context, decisions, and boundaries across sessions — then reasons over that history before the agent acts, steering it away from mistakes. runs locally, we got no access to ur code etc. submitted by /u/JuniorRun3648 [link] [comments]

r/alphaandbetausers8d ago

Coding agents leak tokens across several channels → noisy tool output (test logs, git diff , build spam), model verbosity, thinking tokens, always-loaded instruction files. Existing tools each hit one channel (Caveman → output style, RTK → shell output). HarnessTrim coordinates them behind one cross-harness policy , using the hook/skill primitives the harnesses already expose. Design choices this crowd will care about: Deterministic + idempotent reducers - same input → byte-identical output. No

r/LocalLLaMA8d ago

My understanding is that Claude CLI can be airgapped and told not to touch the internet to avoid hallucinating citations but I haven't gotten to that point in my build yet. We are using Eve and CoCounsel with Westlaw to avoid that right now while I build out some automation with Claude but my goal is to do essentially what you're doing and tie Claude to CoCounsel to be able to do everything in one place. I have a few now outdated cyber security certs so I know enough to be dangerous but

r/legaltech9d ago
Source preview · reddit.com

submitted by /u/fagnerbrack [link] [comments]

reddit.com9d ago

Hey everyone, I just finished porting the Claude Code Game Studio over to CodeX so the AI will hallucinate less just because I ran Claude Code skills using CodeX. All prompts and agents are now run natively to CodeX cli instead of CC If you've tried using AI for game dev, you know standard chats quickly lose context, forget your GDD, and break engine code. Codex Game Studio fixes this by turning your local session into a structured, file-backed studio using plain text files (.toml, .md) trac

r/gamedev9d ago

i use claude code daily and measured what pages cost it while doing research. an average wikipedia article, for instance, is 68,240 tokens of raw html (tiktoken); nike's homepage is 353,000.claude code's built-in webfetch handles the easy case well. it summarizes wikipedia to about 950 tokens and clears cloudflare on some sites like indeed and ticketmaster. but, and there's always a but, on js-rendered and some anti-bot pages it returns nothing.quotes.toscrape.com/js gives "no quotes found"; nik

HN10d ago
Source preview · reddit.com

Not having to manually feed it context every time is a big deal imo

reddit.com10d ago
Source preview · reddit.com

Build Docusign. Make no mistakes.

reddit.com10d ago
Source preview · reddit.com

you're on the pro plan?

reddit.com10d ago
Source preview · reddit.com

Did you use fable?

reddit.com10d ago

You see the absurdity, right? Claude Code is the same rampage companion. Hence, tailored apps with a set of tools will still be in need.

r/legaltech10d ago

Use Claude cowork opus to refactor for you. Back up first! Work with it to make your specifications clear and have it document them as a markdown file. Then have it refactor.

r/ObsidianMD10d ago

Yes. I am using Claude to pull information to create weekly client reports. I do have a disclaimer on those reports that it was created by AI and I do review them before they go out. But it’s saved me a lot of time. I also use it for internal linking suggestions as all my blog posts are saved in a database. I’m sure I will continue to find other uses.

r/Notion10d ago
Source preview · reddit.com

I am not

reddit.com10d ago

We've just started to look at this ourselves. We're looking at a mixture of technical restrictions, written policy, and supported flows (which we'll gradually expand) e.g. Claude/Codex usage must be via Dev Containers or VMs. Technical restrictions include for example: Teams/Enterprise account with SSO enforced. Access limited to compliant corp devices with Passwordless or Phishing resistant auth. Group mappings via IDP Managed Claude enterprise config (managed-settings.json) via Cla

r/sysadmin11d ago

Good question! They solve two different problems. Claude Code’s memory (CLAUDE.md / memory files) stores what’s true right now facts and preferences. Stuff like “we use TypeScript,” “the database is PostgreSQL,” or “user prefers tabs.” It’s a flat fact sheet that prevents simple forgetting. NodeDex stores the reasoning history the chain of decisions, dead ends, constraints, and insights the agent ran into. It answers: · What did we try and abandon, and why? · What alternatives were considered be

r/SideProject12d ago

Which tool are you using? Just standard CLI and/or ollama/openwebui? Or something that is actually geared towards coding, like Roo or Aider/Cline? You'd also have to "pay the context fee", where the agent will read and analyze your code, possibly requiring 100k or more tokens. (For Qwen3.6-27b with 4bit quantization requires about 32GB VRAM)

r/LocalLLaMA12d ago

Try one-context-mcp - free opensource One local MCP server that gives Claude, Cline, Codex, and other AI tools the same project memory. https://github.com/m4vic/one-context-mcp

r/alphaandbetausers12d ago
Source preview · reddit.com

Yeah, reviewing Claude output can feel like a job now.

reddit.com12d ago
Source preview · reddit.com

You dont need context poisoning

reddit.com13d ago

idea is really good, not only for claude/codex instead anything. Browser tabs after full load/ url downloading something, discord channel texts… and you can also give it some extra setting like “screen visual”: could work like an “alt+tab” previewer of sessions/ tabs/… and each state +keybind to switch. more scalability i think? anyways, is ur proyect, looking a solid idea to me ^^

r/SideProject13d ago

Just use Claude Code to supplement with your learning man. A few concepts you should look into as well: Loop, Harness Engineering. You can adjust the output settings to be a "teacher or tutor. Don't forget a bunch of new plugins these days: context7, exa, superpowers, kitstarter.dev , gstack. We learnin' everydayyyyy

r/webdev13d ago

I plan my specs. Claude implement it. 80% of my daily work now is testing Claud's output. submitted by /u/RandomUserName323232 [link] [comments]

r/webdev13d ago

Usage trackers for Claude Code aren't new, but they all live in a terminal or menu bar and I kept forgetting to check them. So I put one in the Dynamic Island and lock screen instead. It's just there in your face: current block's cost, token count, $/hr burn rate, and a countdown to the reset. Backstory: I had Fable credits about to expire and no time to use them, so I one-shot the whole app over a weekend. ActivityKit, gist-sync, dashboard, all mostly working on the first real pass.

r/SideProject14d ago

Hey guys, it would mean a lot to me if you could check out what I made, tbh I got tired of starting sessions from 0 and re-explaining, maintaining md files for my projects. it's not just memory. it remembers across sessions, but on top of that it actually reasons over what it's seen — spots the patterns, and nudges the agent before it repeats a mistake instead of after. I made my life much simpler, so would love to get any feedback on it, since I am not sure whether I am the only weirdo

r/SideProject14d ago

My issue is data sources. Claude in any form only gets what Shopify collects Ended up having to build a separate MCP so I could connect on a deeper level

r/shopify15d ago

I've been using Claude for a lot of my coding lately, and the failure mode I kept hitting wasn't bad output — it was output that looked fine and wasn't. Ask for a fix, get back more changes than expected, everything looks reasonable, then a few days later something subtle breaks and I'm debugging both the original issue and whatever assumptions Claude quietly made along the way. What's actually helped: - Have it explain the current flow before touching anything - Ask directly

r/SideProject15d ago

Looks interesting. This will be pretty useful if it doesn't degrate the LLM performance. I can code more without being locked for 5 hours limit.

r/SideProject15d ago

SOLUTION LANDSCAPE

Brought to you byTop Sectors

A Player feature.See how many ways this pain can be solved, who's already building, and where the gaps are.