Yargila — AI skill for Claude Code
LLM-as-judge halkası — cevapları önce mekanik, sonra nitel olarak yargılar.
How to install Yargila
Installs to ~/.claude/commands/barisguveloglu-bit-kanal-sitesi-yargila.md
mkdir -p ~/.claude/commands && curl -fsSL https://raw.githubusercontent.com/barisguveloglu-bit/kanal-sitesi/HEAD/.claude/commands/yargila.md -o ~/.claude/commands/barisguveloglu-bit-kanal-sitesi-yargila.md Restart Claude Code, or start a new session, for it to be picked up.
What Yargila does
description: LLM-as-judge halkası — cevapları önce mekanik, sonra nitel olarak yargılar argument-hint: [hizli] — "hizli" yazarsan sadece mekanik yargı
Yargıla (cevap kalitesi)
Argüman: **$ARGUMENTS**
`/degerlendir` aramanın doğru yeri bulup bulmadığını ölçer. Bu komut bir adım ötesini ölçer: **doğru parça geldiğinde doğru cevap verildi mi?** Doğru bölümü getirip yine de yanlış okumak, fazla iddia etmek ya da canon'un sustuğu yerde konuşmak mümkün.
Yargıcın kendini yargılaması so
Alternatives in AI
- WeKnora — Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, an 26.9k ★
- Opencodex — Universal provider proxy for OpenAI Codex & Claude Code — use any LLM (Claude, Gemini, Grok, DeepSeek, Ollama… 12.3k ★
- Firecrawl MCP Server — 🔥 Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM c 6.1k ★
Full documentation available on GitHub
View Source RepositoryRelated Skills
Basanite
Vocabulary-tic detector for Claude Code output — frequency drift over your own transcripts, WordNet specificit
LLM Fusion
OpenAI and Anthropic compatible LLM proxy for Ollama Cloud — panel→judge→synth fusion + smart routing, with ta
Eval From Trace
Build an LLM-as-a-Judge eval grounded in real traces from the Progress Observability Platform.
TrajBias
TrajBias: Structural Biases in LLM-as-Judge Evaluation of Agent Trajectories
Evolve Skill
SPIKE — GEPA/DSPy-style offline A/B evolution of ONE prompt-only skill's SKILL.md body against a small fixture
Ground Truth Evals
An LLM eval harness graded by computed ground truth, not an LLM judge. Worked example: poker, served to models
Related Agents
Gozden Gecirici
Kod incelemesi ve güvenlik. İki işi var — (1) kod yazılmadan ÖNCE o iş için güvenlik gereksinimlerini çıkarır
Dogrulayici
Üretilmiş bir işi (rapor, analiz, kod, plan, e-posta) düşmanca gözle denetler ve hataları bulur. Bir işi tesli
Task Plan Verifier
Phase 4 - LLM-as-judge verification of task definitions during planning. Evaluates tasks against spec, strateg