5buyai × AIFUNS
Promoted
← Back to feed
ClaudeDevs
@ClaudeDevs
Red Hat Red Hat MiniMax (official) MiniMax (official) Microsoft Microsoft OpenClaw🦞 OpenClaw🦞 Notion Notion Spline Spline OpenRouter OpenRouter Netflix Netflix Canva Canva Spotify Spotify Twitch Twitch DigitalOcean DigitalOcean Autodesk Autodesk Microsoft 365 Microsoft 365 YouTube YouTube Figma Figma Google Google Business Business Runway Runway Discord Discord GitHub GitHub bolt.new bolt.new Suno Suno ViggleAI ViggleAI OpenArt OpenArt Google DeepMind Google DeepMind ElevenLabs Developers ElevenLabs Developers Steam Steam PomelliByGoogle PomelliByGoogle Fish Audio Fish Audio
All 🖼 Media 🎬 Video
Here's how our team uses Claude Tag for on-call: When an alert fires in Slack, Claude pulls metrics, diffs deploys, and checks flags. It finds a likely cause and proposes a fix, which we can approve and merge. Every minute counts, so we love that it starts right away! https://t.co/15gdEHjI7d
Evals call the model, so they use tokens and results vary. Pilot with `--runs 1` before a full run. Your plugin's hooks and MCP servers run as you, so only evaluate plugins you trust. Run claude update to try it. Docs: https://t.co/K3fKEzO7bI
Then run `claude plugin eval` You'll see each case's score with and without your plugin in your terminal, plus an HTML report with the full detail. If your account supports it, the report is also published as a private artifact.
Start in your plugin's folder and run `claude plugin eval init` You tell Claude what good and bad output looks like and bring a few real prompts. Claude drafts the test cases and checks, pilots the suite, and tells you what a full run will cost.
New in Claude Code: claude plugin eval See what value your plugin is adding, or if it needs more work. You can create test cases, run your plugin or skill against those test cases, score those runs, then run each case again without the plugin to see the differences. https://t.co/qfPU6WHueV
We've also added auto mode to Claude Managed Agents. With `auto`, Claude reviews each tool call based on your intent in `user.message` events and decides whether to run the tool call, deny it, or ask you for input. https://t.co/tCQYx2zyyN
Two fresh updates to Claude Managed Agents: First, we've added a session viewer to the ant CLI. `ant beta:sessions connect` attaches your terminal to a running session, and `--web` opens a web UI served from localhost. https://t.co/1UgRRLyxUA
You can pop out any pane in the Claude Code desktop app into its own window. Drag the diff or terminal to a second screen while Claude keeps working in the main window, then dock it back whenever you want. You can also run sessions side-by-side or stacked. https://t.co/5BF38zm4ky
Aggregated from the public X timeline; copyright belongs to the original author. This site is not affiliated with this account.
✓ Since 2024 ✓ 500+ buyers served ✓ 1,000+ paid orders
💬 Need a hand?
Purchase / payment / account help — chat with us →

Announcements

If you have a credit card, you can register an account on this site, use your credit card to top up your balance in the personal center, and then use the balance to pay for purchasing our products or services.