跳到主要内容
buildradar
登录
主题 · llm-as-judge

llm-as-judge

标记 llm-as-judge 主题、收录中的开源项目,按星标数排序。

项目数
3
总星标数
2,476
平均星标数
825
占比
0.00%

常跟 llm-as-judge 一起出现在同一个项目上的主题。

近期新秀

近 90 天内创建、标记 llm-as-judge 主题的项目。

  • looper@ksimback

    在执行之前,为 Claude Code 设计可视化、需审查的代理循环。

    712
  • ai-engineer-notebooks@calmrocks

    Hands-on, framework-free Colab notebooks for the AI Engineer / Forward Deployed Engineer (FDE) skill set — model APIs, structured output, tool calling, RAG, evals-as-the-spine, agents (loop from scratch, tool design, guardrails, MCP, Skills), fine-tuning vs LoRA, prompt-injection/security, LLMOps, and customer craft. Runs on the free Groq API.

    620
  • Tracely-ai@Jwuthri

    针对 AI agents 的追踪原生 CI/CD — 生产环境失败转化为阻挡 PR 的回归测试。自动检测、分群、封装为隔离案例,在 CI 中以 0 美元重放。

    1,176+0近 7 天星标变化
  • looper@ksimback

    在执行之前,为 Claude Code 设计可视化、需审查的代理循环。

    712+4近 7 天星标变化
  • ai-engineer-notebooks@calmrocks

    Hands-on, framework-free Colab notebooks for the AI Engineer / Forward Deployed Engineer (FDE) skill set — model APIs, structured output, tool calling, RAG, evals-as-the-spine, agents (loop from scratch, tool design, guardrails, MCP, Skills), fine-tuning vs LoRA, prompt-injection/security, LLMOps, and customer craft. Runs on the free Groq API.

    620+25近 7 天星标变化
← 返回主题列表