LLM-Evaluation-s-Always-Fatiguing
Pinned Loading
Repositories
Showing 10 of 18 repositories
- vibe-skills Public
- demo-research-goal-confirmation-agent Public
A demo to demonstrate a dynamic context engineering pattern
- vibe-evaluator Public
- leaf-playground-hub Public
- smolagents Public Forked from huggingface/smolagents
🤗 smolagents: a barebones library for agents. Agents write python code to call tools and orchestrate other agents.
- aider-solver-template Public
- open-webui Public Forked from open-webui/open-webui
User-friendly WebUI for LLMs (Formerly Ollama WebUI)
- leaf-playground Public
A framework to build scenario simulation projects where human and LLM based agents can participant in, with a user-friendly web UI to visualize simulation, support automatically evaluation on agent action level.
Top languages
Loading…
Most used topics
Loading…