10xAI Wiki

此标签下有6条笔记。

  • 2026年8月11日

    Agent Harness(智能体框架/脚手架)

    • agent-harness
    • ai-agent
    • frameworks
    • observability
    • context-engineering
    • llm-eval
  • 2026年8月07日

    LLM 评估

    • llm-eval
    • benchmark
    • llm-as-judge
    • agentops
    • evaluation-platform
    • rag
  • 2026年7月25日

    LLM-HTML 生成

    • llm-html-generation
    • frontend
    • tailwind
    • benchmark
    • llm-eval
    • agent-skills
  • 2026年7月25日

    RAG(检索增强生成)

    • rag
    • retrieval
    • llm-eval
    • multimodal
    • benchmark
    • ai-search
  • 2026年7月25日

    联网搜索评测

    • web-search-eval
    • llm-eval
    • ai-search
    • rag
    • llm-as-judge
    • citation
  • 2026年7月25日

    LLM 评测的主客观分裂与"实验室≠生产"

    • llm-eval
    • benchmark
    • subjective-objective
    • lab-vs-production

Created with Quartz v5.0.0 © 2026

  • RSS
  • GitHub