RamehNagella/mcp_server_fun_eval_llm_judge

Developed an MCP server and implemented an LLM-as-judge evaluation framework to functionally test agent tool-calling behavior, using LangSmith to track experiments and compare results across iterations.

0Stars on the repository
1Mods indexed here, across every type
2mo agoLast push, which is what freshness is scored on
noneNo LICENSE: all rights reserved, so bodies are not copied

mcp-server

01

RamehNagella/mcp_server_fun_eval_llm_judge

MCP server Claude CodeCodexCursor +2

A custom MCP server that provides useful tools and resources for AI assistants. Runs locally from the mcp-server Python package.

not rated 0 2mo ago A tokens not measured