293 tools found
"An Open Source Language Model Specialized in Evaluating Other Language Models."
An open-source visual programming environment for battle-testing prompts to LLMs.
LLM Comparator is an interactive data visualization tool for evaluating and analyzing LLM responses side-by-side, developed by the PAIR team.
The LLM Evaluation Framework
agent to agent communication protocol
research on evaluation of LLMs conducted by Microsoft Research and other collaborated institutes. (Updated at: 2023/10)
"GPTeam uses GPT-4 to create multiple agents who collaborate to achieve predefined goals"
"an experimental open-source attempt to make GPT-4 fully autonomous"
A Challenging, Contamination-Free LLM Benchmark
no code approach to build AI agents
AI multi-agent problem solving
JARVIS, a system to connect LLMs with ML community
Open-Source AI Chatbot / Agent builder with support for LLMs as well as social media channels integration.
summarization of terms related to LLM agents
paper by OpenAI that offers a set of practices for keeping agentsβ operations safe and accountable.
AI agents for insights and research
Agent Framework / shim to use Pydantic with LLMs
A MIT-licensed, deployable starter kit for building and customizing your own version of AI town - a virtual town where AI characters live, chat and socialize.
multi-agent conversation framework as a high-level abstraction by Microsoft [[github](https://github.com/microsoft/autogen)]