@fil_ships If it helps, you can check out a library I wrote specifically for this purpose - trimming context smartly instead of truncation.
https://t.co/hTMpEq7gWc
@maheshboda006 Good to see a side-by-side test. The save step would be interesting to dig into: can annotations made in either build be reopened and edited in Preview? That'd say a lot about file compatibility.
@RABSWFL Nice to see more dictation tools on Linux! Can you skip Claude's cleanup for a single dictation? Keeping filenames and command flags verbatim seems especially useful when talking through code.
@bravish_ Congrats on 1.0! The signed gaps caught my eye. A small replay showing a complete run alongside one with missing events would be a useful demo of what the verifier can and can't tell you.
@KI_Vater Hey Maurice! I'm building open-source tools for AI coding too, including Gauntlet for code review. With Plugin Guard, what's been harder so far: catching risky behavior or keeping false positives manageable?
@zxlzr Yeah, a useful test would be clearing the original task, then seeing whether the agent can recover a definition it needs later without being told to look for it.
@itzKashan2912 Try https://t.co/bLnt6q8P5o in a browser while logged into the affected account. Briefly explain why you think it was flagged by mistake. What happens when you try: an error, a disabled form, or trouble logging in?
One rule in Gauntlet: a review pass that changes the code canβt also count as the clean pass.
Run the relevant checks on the final version, then give the repair a fresh review. A passing check only tells you about the version it ran against.
@bigcodegabo Nice to see this built for Ubuntu. The tool cards showing what Neptune actually executed caught my eye in the docs. That should make it easier to check the lead agentβs summary when something goes wrong.
@omarvvvr For a bug fix, Iβd start with a reproducer that fails before the patch and passes after it. Then run the wider tests and try a nearby edge case. That helps check whether the fix addresses the actual bug.
@b4dgerr Fair point π Are you looking for a coding agent to use day to day, or a framework for building your own agents? That would help narrow the shortlist.