Jev Curate — AI skill for Claude Code
High-throughput synthetic & pretraining dataset sifter powered by TypeSafe AI Jev (api.typesafe.ai).
How to install Jev Curate
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open AkashPriyadarshii/jev-curate and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Jev Curate does
High-throughput synthetic & pretraining dataset sifter powered by TypeSafe AI Jev (api.typesafe.ai). Stream, filter, and score Parquet & JSONL datasets at 1,500+ rows/sec using System One typed decisions (Choice, Score, Noul).
Alternatives in AI
- User Research Skill — Cookiy AI Skill for AI agents (Claude, Codex, Cursor, OpenClaw) — end-to-end user research: AI interviews, syn 1.5k ★
- Graymatter — 30 sec to give your AI agents persistent memory 460 ★
- onWatch — Open-source Go CLI that tracks AI API quota usage across 7 providers (Synthetic, Z.ai, Anthropic, Codex, GitHu 397 ★
README
`jev-curate`
**High-Throughput Synthetic & Pretraining Dataset Sifter Powered by TypeSafe AI (Jev)**
[](https://crates.io/crates/jev-curate) [](https://pypi.org/project/jev-curate/) [](LICENSE) [](https://typesafe.ai)
By **[Akash Priyadarshi](https://github.com/AkashPriyadarshii)**
[Why jev-curate](#why-jev-curate) • [Quickstart](#quickstart) • [CLI Reference](#cli-reference) • [Python API](#python-api) • [Architecture](#architecture) • [Non-Goals](#non-goals) • [Ecosystem](#ecosystem)
Why `jev-curate`?
Cleaning 10M to 1B rows of synthetic reasoning data, instruction tuning pairs, or web-scraped corpora is an economic and technical nightmare:
- Generative LLMs are too slow and expensive: Running Claude 3.5 Sonnet or GPT-4o to judge synthetic rows costs $15,000–$50,000 per
Related Skills
SkillEvaluator
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic
Jev Pruner
Claude Code plugin: trim long Bash output with TypeSafe Jev before the model sees it
Canny
Stops AI coding agents from claiming work is done without evidence. Deterministic hooks decide, TypeSafe's Jev
Jev Search Rerank Eval
Does a TypeSafe Jev rerank beat embedding search? Graded relevance eval (9,831 pairs, 164 zh/en queries) over
Skillranker
Rust CLI powered by Jev from TypeSafe.ai that ranks agent skills for the next step using live session context.
Jev Flash Router
open-sourced jev-flash-router: an MCP server for TypeSafe's new Jev model. AI coding agents waste hundreds of
Related Agents
Data Reviewer
Reviews datasets, schemas, historian tag lists, MES/ERP tables, PI/AF trees, CSV/Parquet exports, Kafka topics
Typesafe Adversary
Red-teams TypeSafe integrations before users do. Probes the nine documented Jev failure modes, tests prompt in
Data Builder
Builds and audits datasets for tiny models — synthetic generation, teacher labelling, held-out splits, and lab