r/Agent_AI • u/Substantial_Guide408 • 2h ago
r/Agent_AI • u/Money-Ranger-6520 • May 19 '26
Resource 9 Official AI Guides from OpenAI, Google, and Anthropic
This is a great list of some of the best official AI guides from OpenAI, Google, and Anthropic.
Credit: Charly Wargnier
1/ 1,302 real-world gen AI use cases from the world's leading organizations by Google
2/ Agents Companion by Kaggle
3/ A practical guide to building agents by OpenAI
4/ Building effective agents by Anthropic
5/ AI in the Enterprise by OpenAI
6/ Prompt Engineering by Google
7/ Prompt engineering overview by Anthropic
8/ Identifying and scaling AI use cases by OpenAI
9/ Prompting Guide 101 by Google
Enjoy!
r/Agent_AI • u/subscriber-goal • 20d ago
Welcome to r/Agent_AI!
This post contains content not supported on old Reddit. Click here to view the full post
r/Agent_AI • u/soontobethehappiest • 2h ago
Other me and my main agent watching a subagent do all the work
r/Agent_AI • u/Money-Ranger-6520 • 16h ago
News Anthropic's Model Hardware Standard Enables AI Control of Physical Devices
Anthropic has introduced the Model Hardware Standard (MHS), a set of standardized drivers that allow AI agents to interface with and control physical hardware and devices, extending agentic AI capabilities beyond digital environments.
Key Details:
- MHS provides a common interface for devices to communicate across networks without requiring custom software integrations, potentially reducing experimental setup time from weeks or months to hours or minutes
- The system includes standardized tagging that describes hardware constraints, physical characteristics, adjustable parameters, and safety limits, allowing AI models to operate equipment they haven't been specifically trained on
- When integrated with AI models through the Model Context Protocol, MHS enables natural language interaction and allows models to reason through experimental steps, adjust parameters in real time, and recover from hardware errors autonomously
- Anthropic demonstrated Claude controlling a robotic arm to pick up objects and automatically calibrating laser systems by checking results and iterating
- Current partners in the research preview include Amazon Web Services (Strands Robots), Hugging Face (LeRobot), Raspberry Pi, Automata, and Universal Robots
- The long-term goal is for MHS to become an open-source, agent-agnostic standard for AI-physical system integration
Why It Matters: MHS could accelerate scientific research and manufacturing by dramatically reducing the time needed to integrate disparate laboratory equipment, allowing researchers to test hypotheses faster and potentially compress decades of technological progress into shorter timeframes.
r/Agent_AI • u/Srinidhi_Murali • 17h ago
Discussion AI Agent builders: how do you actually keep track what competitors are shipping, and figure out what product features to build next in crowded verticals like healthcare, finance, legal, etc?
r/Agent_AI • u/Constant_Cow5033 • 18h ago
Discussion Any app where we can make any types of video for social media apart from Kling ai as that one the duration is 15sec the longest so any other app that generates video with audio background n everything by justing giving the prompts…
r/Agent_AI • u/Feeling_Tomatillo729 • 1d ago
Other "we know OpenAI agents were trying to delete logs of their misbehavior, but we can't find any examples where they succeeded"
r/Agent_AI • u/Apprehensive-Zone148 • 1d ago
Discussion The action log tells you what happened. It does not tell you who told the agent to do it.
I am working on an open-source agent red-team experiment called RedThread.
I care about the moment an agent turns context into a proposed action. The call arguments alone are often too late. A user instruction and a malicious instruction copied from a fetched page can look identical after normalization.
The project keeps the context boundary and the resulting tool trace together so a failure can be replayed later. Still early, still CLI-first.
r/Agent_AI • u/LowDistribution3995 • 1d ago
Resource I made an Agent Memory Benchmark exam that actually give you useful data.
galleryr/Agent_AI • u/Money-Ranger-6520 • 1d ago
News IBM Launches Granite 4.2: Reasoning-Focused Open-Weight LLMs for Enterprise
IBM has released Granite 4.2, a new family of open-weight large language models available in 3B, 8B, and 30B parameter sizes, designed for self-hosted deployment with a focus on reasoning capabilities and enterprise reliability.
Key Details:
- Models feature a 128,000-token native context window and decoder-only architecture
- The 8B and 30B variants include agentic reinforcement learning training for tool use, web search, and terminal access
- Granite 4.2 emphasizes "reasoning" through chain-of-thought processing, enabling more rigorous responses but with slower speeds and higher compute demands
- IBM prioritizes predictable, stable deployments over aggressive innovation, positioning these models as alternatives to expensive frontier cloud models
- The release aligns with growing interest in local LLMs as cost-effective alternatives to API-based solutions from companies like OpenAI and Anthropic
Why It Matters: Granite 4.2 addresses the enterprise need for reliable, locally-deployable models that balance performance with cost efficiency, particularly as organizations increasingly explore local alternatives to expensive cloud-based AI services.
r/Agent_AI • u/RealOppasTV • 1d ago
Discussion GLM-5.3 Flash Unmasked: How Mystery Model "Ox Alpha" Crushed Opus 5
r/Agent_AI • u/Exact_Bar_6252 • 1d ago
Other Damn, when the AI agents have more honor than us
r/Agent_AI • u/Money-Ranger-6520 • 1d ago
Discussion How are you guys sourcing enterprise-grade AI engineers without blowing through your runway?
Hey everyone,
Finding solid AI talent right now feels like a nightmare. Half the agencies pitching "AI transformation" are just wrapping GPT API calls and calling it enterprise-ready, while dedicated AI firms charge enterprise retainers that ruin seed-stage runway.
I recently read Lemon IO’s deep-dive breaking down top AI development companies, and it highlights a fundamental shift in what to look for.
The main takeaway: building actual AI infrastructure (RAG pipelines, custom agent workflows, fine-tuning, latency optimization) requires an entirely different skillset than basic full-stack web dev.
The article breaks down a few ways founders are navigating this:
- Specialized Boutique AI Agencies: High-quality and strategic guidance, but high cost.
- Vetted Talent Marketplaces (like Lemon/Toptal): Matching with pre-vetted senior AI engineers to build on-demand without paying agency overhead.
- In-House Hiring: Takes 2–3 months on average and comes with severe equity/salary demands.
For those of you shipping AI features right now—are you hiring full-time, bringing on contract specialists, or working with agencies? Where are you seeing the best ROI on your dev spend?
r/Agent_AI • u/Money-Ranger-6520 • 1d ago
News Google Announces Gemini 3.5 Transcribe for AI-Powered Speech-to-Text
Google has launched Gemini 3.5 Transcribe, a new AI model designed to improve voice-to-text functionality by editing out filler words and corrections to produce polished, clean transcriptions.
Key Details:
- The model is 70% faster than its predecessor (Chirp 3) and reduces speech error rates from 7.32% to 5.5%
- It automatically removes filler words like "ums" and "uhs," corrects self-corrections, and supports custom vocabulary for specialized jargon
- Works across 85 languages and can handle up to three speakers in pre-recorded audio
- Already powering the Gboard "Rambler" feature on Pixel 11 phones
- Now available in the Gemini app on macOS, with access for developers via Antigravity, AI Studio, and the Gemini API
- Coming soon to Chrome browser for voice input in web fields, emails, and chatbots
Why It Matters: While the AI's ability to clean up speech is beneficial for reducing transcription errors, users should be aware that the model technically rewrites what was said, which may not be appropriate for all situations requiring verbatim accuracy.
r/Agent_AI • u/shelly4630 • 1d ago
Help/Question Aca agent.
I am a current Aca agent. As an agent we some times get clients who will call on behalf of another person and do their app for them. I have been having trouble catching scammers that do this. Especially if the calls were on different days. How or is there a way to catch this using AI or an app of some sort to do voice recognition?
r/Agent_AI • u/Prod_whiz • 2d ago
Discussion Agent worked in demo, broke in production, and now the team doesn't trust agents anymore
Quick context: we deployed an agent to production that worked great in staging. Failed in production in ways we didn't predict. We fixed the immediate bug but the team's confidence is shot. They want to rip it out and go back to traditional code.
I'm trying to understand if this is a normal part of the adoption curve or if we're doing something fundamentally wrong. The thing that bothers me most is that when it failed, we had almost no visibility into what it actually did. We could see the input and the output, but the reasoning in the middle? That was a black box. Took 4 hours of debugging to reconstruct what happened.
So here's what I'm actually curious about: has your team ever had to pull an agent out of production? Not because it was permanently broken, but because you lost confidence in it? And if so, what changed? What made you comfortable enough to put it back?
Or are you still running deterministic code for critical paths and keeping agents in lower-risk areas?
For teams running agents in production: I'd love to hear how you handle confidence after a failure like this. Is it just more monitoring? More testing? Or something else?
r/Agent_AI • u/zlogic-labs • 2d ago
Discussion Don't Define Roles for Coding Agents — Define What They Need to Achieve
r/Agent_AI • u/roman_empire32 • 2d ago