Creating a ReAct Agent from the scratch with nodeJS ( wikipedia search )
Introduction
We'll create an AI agent capable of searching Wikipedia and answering questions based on the information it finds. This ReAct (Reason and Act) Agent uses the Google Generative AI API to process queries and generate responses. Our agent will be able to:
- Search Wikipedia for relevant information.
- Extract specific sections from Wikipedia pages.
- Reason about the information gathered and formulate answers.
[2] What is a ReAct Agent?
A ReAct Agent is a specific type of agent that follows a Reflection-Action cycle. It reflects on the current task, based on available information and actions it can perform, and then decides which action to take or whether to conclude the task.
[3] Planning the Agent
3.1 Required Tools
- Node.js
- Axios library for HTTP requests
- Google Generative AI API (gemini-1.5-flash)
- Wikipedia API
3.2 Agent Structure
Our ReAct Agent will have three main states:
- THOUGHT (Reflection)
- ACTION (Execution)
- ANSWER (Response)
[4] Implementing the Agent
Let's build the ReAct Agent step by step, highlighting each state.
4.1 Initial Setup
First, set up the project and install dependencies:
mkdir react-agent-project cd react-agent-project npm init -y npm install axios dotenv @google/generative-ai
Create a .env file at the project's root:
GOOGLE_AI_API_KEY=your_api_key_here
4.2 Creating the Tools.js File
Create Tools.js with the following content:
const axios = require("axios"); class Tools { static async wikipedia(q) { try { const response = await axios.get("https://en.wikipedia.org/w/api.php", { params: { action: "query", list: "search", srsearch: q, srwhat: "text", format: "json", srlimit: 4, }, }); const results = await Promise.all( response.data.query.search.map(async (searchResult) => { const sectionResponse = await axios.get( "https://en.wikipedia.org/w/api.php", { params: { action: "parse", pageid: searchResult.pageid, prop: "sections", format: "json", }, }, ); const sections = Object.values( sectionResponse.data.parse.sections, ).map((section) => `${section.index}, ${section.line}`); return { pageTitle: searchResult.title, snippet: searchResult.snippet, pageId: searchResult.pageid, sections: sections, }; }), ); return results .map( (result) => `Snippet: ${result.snippet}\nPageId: ${result.pageId}\nSections: ${JSON.stringify(result.sections)}`, ) .join("\n\n"); } catch (error) { console.error("Error fetching from Wikipedia:", error); return "Error fetching data from Wikipedia"; } } static async wikipedia_with_pageId(pageId, sectionId) { if (sectionId) { const response = await axios.get("https://en.wikipedia.org/w/api.php", { params: { action: "parse", format: "json", pageid: parseInt(pageId), prop: "wikitext", section: parseInt(sectionId), disabletoc: 1, }, }); return Object.values(response.data.parse?.wikitext ?? {})[0]?.substring( 0, 25000, ); } else { const response = await axios.get("https://en.wikipedia.org/w/api.php", { params: { action: "query", pageids: parseInt(pageId), prop: "extracts", exintro: true, explaintext: true, format: "json", }, }); return Object.values(response.data?.query.pages)[0]?.extract; } } } module.exports = Tools;
4.3 Creating the ReactAgent.js File
Create ReactAgent.js with the following content:
require("dotenv").config(); const { GoogleGenerativeAI } = require("@google/generative-ai"); const Tools = require("./Tools"); const genAI = new GoogleGenerativeAI(process.env.GOOGLE_AI_API_KEY); class ReActAgent { constructor(query, functions) { this.query = query; this.functions = new Set(functions); this.state = "THOUGHT"; this._history = []; this.model = genAI.getGenerativeModel({ model: "gemini-1.5-flash", temperature: 2, }); } get history() { return this._history; } pushHistory(value) { this._history.push(`\n ${value}`); } async run() { this.pushHistory(`**Task: ${this.query} **`); try { return await this.step(); } catch (e) { if (e.message.includes("exhausted")) { return "Sorry, I'm exhausted, I can't process your request anymore. ><"; } return "Unable to process your request, please try again? ><"; } } async step() { const colors = { reset: "\x1b[0m", yellow: "\x1b[33m", red: "\x1b[31m", cyan: "\x1b[36m", }; console.log("===================================="); console.log( `Next Movement: ${ this.state === "THOUGHT" ? colors.yellow : this.state === "ACTION" ? colors.red : this.state === "ANSWER" ? colors.cyan : colors.reset }${this.state}${colors.reset}`, ); console.log(`Last Movement: ${this.history[this.history.length - 1]}`); console.log("===================================="); switch (this.state) { case "THOUGHT": await this.thought(); break; case "ACTION": await this.action(); break; case "ANSWER": await this.answer(); break; } } async promptModel(prompt) { const result = await this.model.generateContent(prompt); const response = await result.response; return response.text(); } async thought() { const availableFunctions = JSON.stringify(Array.from(this.functions)); const historyContext = this.history.join("\n"); const prompt = `Your task to FullFill ${this.query}. Context contains all the reflection you made so far and the ActionResult you collected. AvailableActions are functions you can call whenever you need more data. Context: "${historyContext}" << AvailableActions: "${availableFunctions}" << Task: "${this.query}" << Reflect uppon Your Task using Context, ActionResult and AvailableActions to find your next_step. print your next_step with a Thought or FullFill Your Task `; const thought = await this.promptModel(prompt); this.pushHistory(`\n **${thought.trim()}**`); if ( thought.toLowerCase().includes("fullfill") || thought.toLowerCase().includes("fulfill") ) { this.state = "ANSWER"; return await this.step(); } this.state = "ACTION"; return await this.step(); } async action() { const action = await this.decideAction(); this.pushHistory(`** Action: ${action} **`); const result = await this.executeFunctionCall(action); this.pushHistory(`** ActionResult: ${result} **`); this.state = "THOUGHT"; return await this.step(); } async decideAction() { const availableFunctions = JSON.stringify(Array.from(this.functions)); const historyContext = this.history; const prompt = `Reflect uppon the Thought, Query and AvailableActions ${historyContext[historyContext.length - 2]} Thought <<< ${historyContext[historyContext.length - 1]} Query: "${this.query}" AvailableActions: ${availableFunctions} output only the function,parametervalues separated by a comma. For example: "wikipedia,ronaldinho gaucho, 1450"`; const decision = await this.promptModel(prompt); return `${decision.replace(/`/g, "").trim()}`; } async executeFunctionCall(functionCall) { const [functionName, ...args] = functionCall.split(","); const func = Tools[functionName.trim()]; if (func) { return await func.call(null, ...args); } throw new Error(`Function ${functionName} not found`); } async answer() { const historyContext = this.history; const prompt = `Based on the following context, provide a complete, detailed and descriptive formated answer for the Following Task: ${this.query} . Context: ${historyContext} Task: "${this.query}"`; const finalAnswer = await this.promptModel(prompt); this.history.push(`Answer: ${this.finalAnswer}`); console.log("WE WILL ANSWER >>>>>>>", finalAnswer); return finalAnswer; } } module.exports = ReActAgent;
4.4 Running the agent (index.js)
Create index.js with the following content:
const ReActAgent = require("./ReactAgent.js"); async function main() { const query = "What does England border with?"; const functions = [ [ "wikipedia", "params: query", "Semantic Search Wikipedia API for snippets, pageIds and sectionIds >> \n ex: Date brazil has been colonized? \n Brazil was colonized at 1500, pageId, sections : []", ], [ "wikipedia_with_pageId", "params : pageId, sectionId", "Search Wikipedia API for data using a pageId and a sectionIndex as params. \n ex: 1500, 1234 \n Section information about blablalbal", ], ]; const agent = new ReActAgent(query, functions); try { const result = await agent.run(); console.log("THE AGENT RETURN THE FOLLOWING >>>", result); } catch (e) { console.log("FAILED TO RUN T.T", e); } } main().catch(console.error);
[5] How the Wikipedia Part Works
The interaction with Wikipedia is done in two main steps:
-
Initial search (wikipedia function):
- Makes a request to the Wikipedia search API.
- Returns up to 4 relevant results for the query.
- For each result, it fetches the sections of the page.
-
Detailed search (wikipedia_with_pageId function):
- Uses the page ID and section ID to fetch specific content.
- Returns the text of the requested section.
This process allows the agent to first get an overview of topics related to the query and then dive deeper into specific sections as needed.
[6] Execution Flow Example
- The user asks a question.
- The agent enters the THOUGHT state and reflects on the question.
- It decides to search Wikipedia and enters the ACTION state.
- Executes the wikipedia function and obtains results.
- Returns to the THOUGHT state to reflect on the results.
- May decide to search for more details or a different approach.
- Repeats the THOUGHT and ACTION cycle as necessary.
- When it has sufficient information, it enters the ANSWER state.
- Generates a final answer based on all the information collected.
- Enters infinite loop whenever the wikipedia doesn't have the data to collect. Fix it with a timer =P
[7] Final Considerations
- The modular structure allows for easy addition of new tools or APIs.
- It's important to implement error handling and time/iteration limits to avoid infinite loops or excessive resource use.
- Use Temperature : 99999 lol
以上是Creating a ReAct Agent from the scratch with nodeJS ( wikipedia search )的详细内容。更多信息请关注PHP中文网其他相关文章!

热AI工具

Undresser.AI Undress
人工智能驱动的应用程序,用于创建逼真的裸体照片

AI Clothes Remover
用于从照片中去除衣服的在线人工智能工具。

Undress AI Tool
免费脱衣服图片

Clothoff.io
AI脱衣机

Video Face Swap
使用我们完全免费的人工智能换脸工具轻松在任何视频中换脸!

热门文章

热工具

记事本++7.3.1
好用且免费的代码编辑器

SublimeText3汉化版
中文版,非常好用

禅工作室 13.0.1
功能强大的PHP集成开发环境

Dreamweaver CS6
视觉化网页开发工具

SublimeText3 Mac版
神级代码编辑软件(SublimeText3)

JavaScript是现代Web开发的基石,它的主要功能包括事件驱动编程、动态内容生成和异步编程。1)事件驱动编程允许网页根据用户操作动态变化。2)动态内容生成使得页面内容可以根据条件调整。3)异步编程确保用户界面不被阻塞。JavaScript广泛应用于网页交互、单页面应用和服务器端开发,极大地提升了用户体验和跨平台开发的灵活性。

JavaScript的最新趋势包括TypeScript的崛起、现代框架和库的流行以及WebAssembly的应用。未来前景涵盖更强大的类型系统、服务器端JavaScript的发展、人工智能和机器学习的扩展以及物联网和边缘计算的潜力。

不同JavaScript引擎在解析和执行JavaScript代码时,效果会有所不同,因为每个引擎的实现原理和优化策略各有差异。1.词法分析:将源码转换为词法单元。2.语法分析:生成抽象语法树。3.优化和编译:通过JIT编译器生成机器码。4.执行:运行机器码。V8引擎通过即时编译和隐藏类优化,SpiderMonkey使用类型推断系统,导致在相同代码上的性能表现不同。

Python更适合初学者,学习曲线平缓,语法简洁;JavaScript适合前端开发,学习曲线较陡,语法灵活。1.Python语法直观,适用于数据科学和后端开发。2.JavaScript灵活,广泛用于前端和服务器端编程。

JavaScript是现代Web开发的核心语言,因其多样性和灵活性而广泛应用。1)前端开发:通过DOM操作和现代框架(如React、Vue.js、Angular)构建动态网页和单页面应用。2)服务器端开发:Node.js利用非阻塞I/O模型处理高并发和实时应用。3)移动和桌面应用开发:通过ReactNative和Electron实现跨平台开发,提高开发效率。

本文展示了与许可证确保的后端的前端集成,并使用Next.js构建功能性Edtech SaaS应用程序。 前端获取用户权限以控制UI的可见性并确保API要求遵守角色库

我使用您的日常技术工具构建了功能性的多租户SaaS应用程序(一个Edtech应用程序),您可以做同样的事情。 首先,什么是多租户SaaS应用程序? 多租户SaaS应用程序可让您从唱歌中为多个客户提供服务

从C/C 转向JavaScript需要适应动态类型、垃圾回收和异步编程等特点。1)C/C 是静态类型语言,需手动管理内存,而JavaScript是动态类型,垃圾回收自动处理。2)C/C 需编译成机器码,JavaScript则为解释型语言。3)JavaScript引入闭包、原型链和Promise等概念,增强了灵活性和异步编程能力。
