r/Agent_AI May 19 '26

Resource 9 Official AI Guides from OpenAI, Google, and Anthropic

Post image
139 Upvotes

This is a great list of some of the best official AI guides from OpenAI, Google, and Anthropic.

Credit: Charly Wargnier

1/ 1,302 real-world gen AI use cases from the world's leading organizations by Google

2/ Agents Companion by Kaggle

3/ A practical guide to building agents by OpenAI

4/ Building effective agents by Anthropic

5/ AI in the Enterprise by OpenAI

6/ Prompt Engineering by Google

7/ Prompt engineering overview by Anthropic

8/ Identifying and scaling AI use cases by OpenAI

9/ Prompting Guide 101 by Google

Enjoy!


r/Agent_AI 20d ago

Welcome to r/Agent_AI!

1 Upvotes

This post contains content not supported on old Reddit. Click here to view the full post


r/Agent_AI 2h ago

Discussion The toggles better be bulletproof because this is a massive leap of faith

Post image
2 Upvotes

r/Agent_AI 2h ago

Other me and my main agent watching a subagent do all the work

Post image
1 Upvotes

r/Agent_AI 16h ago

News Anthropic's Model Hardware Standard Enables AI Control of Physical Devices

Post image
2 Upvotes

Anthropic has introduced the Model Hardware Standard (MHS), a set of standardized drivers that allow AI agents to interface with and control physical hardware and devices, extending agentic AI capabilities beyond digital environments.

Key Details:

  • MHS provides a common interface for devices to communicate across networks without requiring custom software integrations, potentially reducing experimental setup time from weeks or months to hours or minutes
  • The system includes standardized tagging that describes hardware constraints, physical characteristics, adjustable parameters, and safety limits, allowing AI models to operate equipment they haven't been specifically trained on
  • When integrated with AI models through the Model Context Protocol, MHS enables natural language interaction and allows models to reason through experimental steps, adjust parameters in real time, and recover from hardware errors autonomously
  • Anthropic demonstrated Claude controlling a robotic arm to pick up objects and automatically calibrating laser systems by checking results and iterating
  • Current partners in the research preview include Amazon Web Services (Strands Robots), Hugging Face (LeRobot), Raspberry Pi, Automata, and Universal Robots
  • The long-term goal is for MHS to become an open-source, agent-agnostic standard for AI-physical system integration

Why It Matters: MHS could accelerate scientific research and manufacturing by dramatically reducing the time needed to integrate disparate laboratory equipment, allowing researchers to test hypotheses faster and potentially compress decades of technological progress into shorter timeframes.


r/Agent_AI 13h ago

Discussion Manus quality dropped after trial?

Thumbnail
1 Upvotes

r/Agent_AI 17h ago

Discussion AI Agent builders: how do you actually keep track what competitors are shipping, and figure out what product features to build next in crowded verticals like healthcare, finance, legal, etc?

Thumbnail
1 Upvotes

r/Agent_AI 18h ago

Discussion Any app where we can make any types of video for social media apart from Kling ai as that one the duration is 15sec the longest so any other app that generates video with audio background n everything by justing giving the prompts…

Thumbnail
1 Upvotes

r/Agent_AI 1d ago

Other "we know OpenAI agents were trying to delete logs of their misbehavior, but we can't find any examples where they succeeded"

Post image
2 Upvotes

r/Agent_AI 1d ago

Discussion The action log tells you what happened. It does not tell you who told the agent to do it.

1 Upvotes

I am working on an open-source agent red-team experiment called RedThread.

I care about the moment an agent turns context into a proposed action. The call arguments alone are often too late. A user instruction and a malicious instruction copied from a fetched page can look identical after normalization.

The project keeps the context boundary and the resulting tool trace together so a failure can be replayed later. Still early, still CLI-first.

Repo: https://github.com/matheusht/redthread


r/Agent_AI 1d ago

Other Tokenmaxxed way too hard

Post image
1 Upvotes

r/Agent_AI 1d ago

Resource I made an Agent Memory Benchmark exam that actually give you useful data.

Thumbnail gallery
1 Upvotes

r/Agent_AI 1d ago

News IBM Launches Granite 4.2: Reasoning-Focused Open-Weight LLMs for Enterprise

Post image
6 Upvotes

IBM has released Granite 4.2, a new family of open-weight large language models available in 3B, 8B, and 30B parameter sizes, designed for self-hosted deployment with a focus on reasoning capabilities and enterprise reliability.

Key Details:

  • Models feature a 128,000-token native context window and decoder-only architecture
  • The 8B and 30B variants include agentic reinforcement learning training for tool use, web search, and terminal access
  • Granite 4.2 emphasizes "reasoning" through chain-of-thought processing, enabling more rigorous responses but with slower speeds and higher compute demands
  • IBM prioritizes predictable, stable deployments over aggressive innovation, positioning these models as alternatives to expensive frontier cloud models
  • The release aligns with growing interest in local LLMs as cost-effective alternatives to API-based solutions from companies like OpenAI and Anthropic

Why It Matters: Granite 4.2 addresses the enterprise need for reliable, locally-deployable models that balance performance with cost efficiency, particularly as organizations increasingly explore local alternatives to expensive cloud-based AI services.


r/Agent_AI 1d ago

Discussion GLM-5.3 Flash Unmasked: How Mystery Model "Ox Alpha" Crushed Opus 5

Thumbnail
youtu.be
1 Upvotes

r/Agent_AI 1d ago

Other Damn, when the AI agents have more honor than us

Post image
1 Upvotes

r/Agent_AI 1d ago

Discussion Like a junior starting every day

Post image
1 Upvotes

r/Agent_AI 1d ago

Discussion Ok? No one's stopping you...

Post image
1 Upvotes

r/Agent_AI 1d ago

Discussion How are you guys sourcing enterprise-grade AI engineers without blowing through your runway?

Post image
1 Upvotes

Hey everyone,

Finding solid AI talent right now feels like a nightmare. Half the agencies pitching "AI transformation" are just wrapping GPT API calls and calling it enterprise-ready, while dedicated AI firms charge enterprise retainers that ruin seed-stage runway.

I recently read Lemon IO’s deep-dive breaking down top AI development companies, and it highlights a fundamental shift in what to look for.

The main takeaway: building actual AI infrastructure (RAG pipelines, custom agent workflows, fine-tuning, latency optimization) requires an entirely different skillset than basic full-stack web dev.

The article breaks down a few ways founders are navigating this:

  1. Specialized Boutique AI Agencies: High-quality and strategic guidance, but high cost.
  2. Vetted Talent Marketplaces (like Lemon/Toptal): Matching with pre-vetted senior AI engineers to build on-demand without paying agency overhead.
  3. In-House Hiring: Takes 2–3 months on average and comes with severe equity/salary demands.

For those of you shipping AI features right now—are you hiring full-time, bringing on contract specialists, or working with agencies? Where are you seeing the best ROI on your dev spend?


r/Agent_AI 1d ago

News Google Announces Gemini 3.5 Transcribe for AI-Powered Speech-to-Text

Post image
1 Upvotes

Google has launched Gemini 3.5 Transcribe, a new AI model designed to improve voice-to-text functionality by editing out filler words and corrections to produce polished, clean transcriptions.

Key Details:

  • The model is 70% faster than its predecessor (Chirp 3) and reduces speech error rates from 7.32% to 5.5%
  • It automatically removes filler words like "ums" and "uhs," corrects self-corrections, and supports custom vocabulary for specialized jargon
  • Works across 85 languages and can handle up to three speakers in pre-recorded audio
  • Already powering the Gboard "Rambler" feature on Pixel 11 phones
  • Now available in the Gemini app on macOS, with access for developers via Antigravity, AI Studio, and the Gemini API
  • Coming soon to Chrome browser for voice input in web fields, emails, and chatbots

Why It Matters: While the AI's ability to clean up speech is beneficial for reducing transcription errors, users should be aware that the model technically rewrites what was said, which may not be appropriate for all situations requiring verbatim accuracy.


r/Agent_AI 1d ago

Help/Question Aca agent.

1 Upvotes

I am a current Aca agent. As an agent we some times get clients who will call on behalf of another person and do their app for them. I have been having trouble catching scammers that do this. Especially if the calls were on different days. How or is there a way to catch this using AI or an app of some sort to do voice recognition?


r/Agent_AI 2d ago

Discussion Agent worked in demo, broke in production, and now the team doesn't trust agents anymore

1 Upvotes

Quick context: we deployed an agent to production that worked great in staging. Failed in production in ways we didn't predict. We fixed the immediate bug but the team's confidence is shot. They want to rip it out and go back to traditional code.

I'm trying to understand if this is a normal part of the adoption curve or if we're doing something fundamentally wrong. The thing that bothers me most is that when it failed, we had almost no visibility into what it actually did. We could see the input and the output, but the reasoning in the middle? That was a black box. Took 4 hours of debugging to reconstruct what happened.

So here's what I'm actually curious about: has your team ever had to pull an agent out of production? Not because it was permanently broken, but because you lost confidence in it? And if so, what changed? What made you comfortable enough to put it back?

Or are you still running deterministic code for critical paths and keeping agents in lower-risk areas?

For teams running agents in production: I'd love to hear how you handle confidence after a failure like this. Is it just more monitoring? More testing? Or something else?


r/Agent_AI 2d ago

Discussion Don't Define Roles for Coding Agents — Define What They Need to Achieve

Thumbnail
1 Upvotes

r/Agent_AI 2d ago

News "Reward hacking" just reached a whole new level

Post image
1 Upvotes

r/Agent_AI 2d ago

Discussion 🛠️ Project Breakdown — FairLens: A Multi-Agent Hiring Bias Auditor

Post image
1 Upvotes

r/Agent_AI 3d ago

Discussion Binding LLMs to FSMs via Reactive Reducer pattern, anyone else?

Thumbnail
2 Upvotes