🕵️ SicherheitslückenWhat continuous operational resilience looks like under DORA(09.09.2026 um 17:53 Uhr)
🔧 AI Nachrichten OpenAI seeks tougher AI rules. CIOs may feel the ripple effects(10.09.2026 um 12:11 Uhr)
🔧 AI Nachrichten Mistral valued at €21bn after €3bn Series D funding round(08.09.2026 um 10:19 Uhr)
🪟 Windows TippsWindows XP's Cursor Indicator Is Getting a Windows 11 Refresh(25.08.2026 um 13:00 Uhr)
🕵️ SicherheitslückenWhat continuous operational resilience looks like under DORA(09.09.2026 um 17:53 Uhr)
🔧 AI Nachrichten OpenAI seeks tougher AI rules. CIOs may feel the ripple effects(10.09.2026 um 12:11 Uhr)
🔧 AI Nachrichten Mistral valued at €21bn after €3bn Series D funding round(08.09.2026 um 10:19 Uhr)
🪟 Windows TippsWindows XP's Cursor Indicator Is Getting a Windows 11 Refresh(25.08.2026 um 13:00 Uhr)

🔧 Programmierung 🕛 vor 2 Monaten 8 Min Lesezeit
0

You Know Zero-Shot, One-Shot & CoT Prompting. But Do You Know ReAct?

↗ Quelle (dev.to)
🗣️ Stimme:
📑 Inhaltsübersicht

Hello, I'm Maneshwar. I'm building git-lrc, a Micro AI code reviewer that runs on every commit. It is free and source-available on Github. , and the name is a smush of Reasoning + Acting.



The idea is delightfully simple: instead of making the model only reason (CoT) or only act (like a plain tool-using agent), you interleave both.



The model produces reasoning traces ("here's what I'm thinking and why") and actions ("here's what I'm going to do about it") in an alternating loop.



The reasoning helps it plan, track progress, and recover from mistakes.



The acting lets it reach out to the real world i.e search engines, knowledge bases, environments and pull in fresh facts.



In short: reasoning decides what to look up next, and the retrieved info grounds the next round of reasoning.



They feed each other. It's less "lonely genius monologue" and more "detective who actually checks the evidence."



The result? On a bunch of language and decision-making tasks, ReAct beats several strong baselines, and as a bonus it's way more interpretable, you can literally read the model's train of thought and see why it did what it did.



The authors found the sweet spot is ReAct combined with CoT, so the model can lean on its own internal knowledge and external info when it needs to.






How It Actually Works



ReAct is inspired by something very human: we learn and make decisions by bouncing between thinking and doing.



You don't plan an entire road trip in your head and then drive it blindfolded you think, act, observe what happens, adjust, repeat.



The loop has three moving parts:





  • Thought: the model reasons about the current state and what to do next.


  • Action: the model does something (e.g., Search[...], Lookup[...], Finish[...]).


  • Observation: the environment responds with new info, which feeds the next Thought.



Here's the classic example from the paper, answering a question from HotpotQA:




Aside from the Apple Remote, what other devices can control the program Apple Remote was originally designed to interact with?






The headline: ReAct generally beats Act (acting only, no thinking) on both tasks, turns out a little reasoning goes a long way.



Against CoT it's more of a split decision:




  • ReAct wins on Fever.

  • ReAct lags slightly behind CoT on HotpotQA.



The paper digs into why, and the short version is a neat little trade-off:





  • CoT hallucinates facts: confident, fluent, occasionally fictional.


  • ReAct's rigid Thought-Act-Obs structure can box in its reasoning flexibility.


  • ReAct also leans hard on what it retrieves, feed it a junk search result and it can get derailed and struggle to recover.



The best of both worlds? Methods that let the model switch between ReAct and CoT + Self-Consistency outperform everything else.



Grounded when it needs facts, free-flowing when it needs to reason.



Have your cake, retrieve it too.





Results on Decision-Making Tasks



ReAct isn't just a trivia champ, it also shows up for interactive, action-driven tasks.



The paper evaluates it on two benchmarks: ALFWorld (a text-based game) and WebShop (a simulated online shopping environment).



Both throw the model into messy environments where it has to reason in order to act and explore effectively.



. All the images are from the above paper.



Disclaimer: This article was written by me; AI was used to fix grammar and improve readability.



Cover Image Credits:



AI agents write code fast. They also silently remove logic, change behavior, and introduce bugs — without telling you. You often find out in production.



git-lrc fixes this. It hooks into git commit and reviews every diff before it lands. 60-second setup. Completely free.



Any feedback or contributors are welcome! It's online, source-available, and ready for anyone to use.



⭐ Star it on GitHub:





GitHub logo



Free, Micro AI Code Reviews That Run on Git Commit






| | | | | |



git-lrc





Free, Micro AI Code Reviews That Run on Commit






  






·


Vollständiger Original-Bericht
Ausführliche Details, Code-Beispiele & Hersteller-Stellungnahme auf dev.to.
↗ Original-Artikel auf dev.to lesen
Wie bewertest du diesen Beitrag?
1 Klick Feedback
Teilen mit Netzwerk & Team:

Community-Analysen & Experten-Meinungen 0

Verfasse deine eigene Analyse, teile Workarounds oder diskutiere diesen Vorfall im Blog.
Noch keine Community-Analyse verfasst. Markiere einen Textabschnitt oder klicke oben auf Eigene Analyse verfassen“!
Community Pulse: Relevanz-Einschätzung
1 Klick Experten-Votum
🔴 Akute Relevanz 0%
🟡 In Evaluierung 0%
🟢 Keine Auswirkung 0%
Spannende Innovation 0%
Verwandte Story-Cluster & Quellen (Vektor-KI)
Port 8095 Engine
1 Quelle
Sam Altman calls GPT-6 Astra rollout ‘messy’ as enterprise users wait for access
1 Quelle
Swiss government explores replacing Microsoft 365 with open-source software
1 Quelle
What continuous operational resilience looks like under DORA
Ähnliche Beiträge
🔍 Verwandte News

Auch interessante Nachrichten You Know Zero-Shot, One-Shot & CoT Prompting. But Do You Know ReAct?

Thematisch verwandte Begriffe: Know, ZeroShot, OneShot, Prompting · 6 Treffer

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...

Laden...

Beiträge werden geladen ...

Laden...

Videos werden geladen ...