- Community Jev-like model and tooling experiment inspired by TypeSafe System One.
vinnylarouge/jevlike · Python
1,283113jevdemomodel - jev-experiments: TypeSafe Jev ecosystem repository.
dabit3/jev-experiments · TypeScript
38128jevtypescript - Camera-only autonomous drone in MuJoCo with a small judgment model (TypeSafe Jev) in the loop at 2.5Hz
RomanSlack/jev-drone · Python
17624jevdemo - 68jevbenchJevBench v1 - a benchmark for Jev-class typed decision models: smart, cheap, fast, reliable, open.
fstandhartinger/jevbench · Python
11812jevpython - Using Jev as an evaluator.
danielgshea/jev-as-a-judge · Python
8514jevpython - 77jevalsAgent evals and guardrails as Jev decisions: one request per trace, a fraction of a cent, fast enough for the agent loop. Runs locally with Kev or Laya.
openlayer-ai/jevals · Python
835jevagentsevalsguardrailsllmllm-evaluationragastypesafe - 92jevfireJEV-inspired parallel decisions for CUDA LLMs. One context, many decisions. vLLM API, game-agent examples, and reproducible benchmarks.
kikoncuo/jevfire · JavaScript
645jevjavascriptcudagame-aiinference - Open reproduction of TypeSafe Jev: a 150M typed decision engine (noul/choice/score in one non-autoregressive pass, calibrated confidence). 0.697 vs Jev's 0.727, 2.5x better calibrated, 4x faster, free. Trains on a Colab T4 in 30 min.
intikhab49/open-jev-typed-decision-engine · Python
432jevagent-observabilitycalibrationcolabdecision-modelexpected-calibration-errorhuggingfacellm-alternative - The highly anticipated open-source repository for JEV as Policy enables one-click setup of the simulation environment. Evaluations of Astra + JEV on benchmarks such as RoboTwin will also be released soon.
YuanKJing/Jev-as-Policy · Python
402jevpython - 108open-jevTyped JSON inference with DiffusionGemma, with Every and Jev benchmark results
JoshuaSP/open-jev · Python
392jevpython - This is a LLM Gateway that mimics typesafe ai structured output. Like an imposter Jev.
iammrduncan/typesafe-ai-benchmark · TypeScript
385jev - Open replica of TypeSafe's Jev: typed calibrated decisions in one forward pass, on Gemma 4 E2B / Gemma 3 270M (Modal)
mithalouni/system-one-open · Python
354jevpython - 115openjevLocal bilingual probability decisions from context, questions, and candidate answers. Independent research preview inspired by TypeSafe Jev.
zhihz/openjev · Python
333jevpython - Independent, evidence-based map of when TypeSafe's Jev actually holds up vs. breaks down — real API-call receipts, not a leaderboard. 中文為主的雙語 repo。
Zaious/jev-capability-atlas · Python
266jevai-agentsbenchmarkcalibrationllm-evaluationmachine-learningtypesafezh-tw - Unofficial study: Jev-style parallel typed decisions on stock 1.5B-8B models on an Apple Silicon laptop. Benchmarks, research notes, and a Hugging Face Space demo.
rorshopping/jev-on-a-laptop · Python
241jevpython - Jev vs Gemini 3.8 Flash: labelling 1,000 app reviews, 4.1× faster and 7× cheaper
goodrahstar/jev-column-race · JavaScript
233jevjavascript - Probability-aware evaluation for typed decision models: calibration, selective risk, latency, and reproducible benchmarks.
AbdelStark/jev-benchmarks · Python
173jevpythonbenchmarkingcalibrationevaluation - 152jevalMeasures what your Jev classifier's confidence is really worth, and sets the human hand-off line from what a mistake costs.
rlaope/jeval · Python
171jevpython - 165jevbetterA stronger one-pass scorer over a variable list of text options. Hashed n-gram encoder, rival-aware attention, gated head, temperature scaling — with a head-to-head benchmark vs the jevlike starter design.
olanotolu/jevbetter · Python
143jevpython - 168jev_stockAn experimental JEV-powered framework for forecasting short-term stock price direction from structured market data.
sosopop/jev_stock · Python
133jevpythonfinancial-datafinancial-machine-learningllm - Interactive experiments with TypeSafe Jev, from support routing to 3D driving simulations with real AI decisions and visible sensor inputs.
kavehmz/typesafe-playground · JavaScript
134jevjavascript - 210trade-jevBacktest Jev (TypeSafe) as a BUY/SELL/HOLD trader on NQ L10 order-book data
justinhe16/trade-jev · Python
83jevpython - 200jev-dspy-labReproducible calibration and selective-risk benchmarks for Jev/TypeSafe decisions in DSPy workflows
jmanhype/jev-dspy-lab · Python
80jevpython - Can a decision model beat dedicated rerankers? TypeSafe Jev vs Cohere Rerank 4 vs ZeroEntropy zerank-2 vs a chat-model baseline: 14 datasets, every raw API response, bootstrap ranges on every gap.
anessbelbati/jev-rerank-bench · Python
70jevpython - 221jev-mcpConnect JEV to MCP clients and compare its judgments against general-purpose LLMs using shared datasets and measurable accuracy.
arunav25/jev-mcp · JavaScript
70jevjavascript - ui-generator-instinct-jev: TypeSafe Jev ecosystem repository.
joevidev/ui-generator-instinct-jev · TypeScript
70jevtypescript - Reproducible early-access evaluation of Jev on Korean understanding and medical text, with runtime and cost evidence
mahlernim/jev-korean-benchmark · Python
60jevpython - 236jev-lmA word-level language model whose output layer is Jev: n-gram drafter, Noul chunk verification, bits-per-token eval
y0usaf/jev-lm · TypeScript
61jevtypescript - Benchmarks and a playground for TypeSafe's Jev (System One) model: chess, and who-is-the-player-talking-to for speech-to-text game NPCs
wondertwins/jev-benchmark · Python
61jevpython - High-speed recursive AI Elo tournament engine powered by Jev and Swiss matchmaking
opaielsheikh/ai-elo-ranker · Python
52jevpython - A show-and-tell capability study for Jev, TypeSafe's System One decision model.
lbotinelly/jev-little-airways · HTML
50jevhtml - Jev (TypeSafe) vs Claude Haiku 4.5 on 2 000 phishing emails: accuracy, calibration, latency, cost. Reproducible benchmark.
anisselbd/jev-phishing-bench · Python
50jevpython - Jev-shaped typed-decision model (state + Choice/Score/Noul questions -> calibrated probabilities, one pass) on ModernBERT / DeBERTa / LLaDA-MoE, with measured latency, accuracy, calibration and training cost
kotoba-lang/typed-decisions · Python
50jev - 252jev-2048An instrumented 2048 web lab where every move is a Jev (TypeSafe AI System One) Choice, with no heuristic fallback | 用 Jev 决策模型驱动每一步的 2048 网页实验台,概率、置信度、延迟与成本全部摊开可见,且刻意不做启发式兜底
ARCJ137442/jev-2048 · TypeScript
51jev2048llm-evaluationsystem-onetyped-decisionstypesafe-aitypescript - 296qwen-rlcdJev-style calibrated decision model (Choice/Score/Noul) on Qwen3.5-0.8B
shamazharikh/qwen-rlcd · Python
41jevpython - 298rubikjevChallenge the Jev's intelligence in Rubik Cube puzzles
0xtrou/rubikjev · TypeScript
40jevtypescript - Independent Jev 1.13.0 behavior study: report, controlled prompt experiments, raw results, and offline verification.
RINNECODER/jev-behavior-study · Python
30jevpython - Blind security benchmarks for Jev, TypeSafe's System One model: prompt injection and vulnerable code detection, built on jev-go
Gaurav-Gosain/jev-sec-bench · Go
30jevgo - A playground for experiments around Jev, TypeSafe's System One model.
markjaquith/typesafe-ai-playground · Rust
31jevrust - Jev (TypeSafe) exploratory thread: claim audit, live demos, and runnable code
SamuelSacco/jev-exploration · Python
30jevpython - Three measured experiments on RAG hallucination: quote-checking, TypeSafe's Jev, and IBM's STAIR. 850+ graded questions, raw responses included.
aryanchauhanoffical/no-hallucination · Python
30jevbenchmarkevaluationhallucinationllmragretrievaltypesafe-ai - About calibrating Jev for code reviews
Selmar/typesafe-jev-calibrate-for-code-review · Python
30jevpython - 333jev-rlJEV Reinforcement Learning: four classic games trained with JEV-powered rewards, reproducible experiments and checkpoint replays.
Bring-AI/jev-rl · Python
30jevpython - 330jev-playsjev-plays: TypeSafe Jev ecosystem repository.
mansicer/jev-plays · Python
30jevpython - 382calibreMeasure when to use Jev and other models on your data, then route accordingly.
FirasSX914/calibre · Python
20jevpythonbenchmarkcalibrationconfidence - Experimental multi-horizon BTC signal generator using TypeSafe Jev probabilities and Binance market data.
WebGrga/btc-jev-signal · TypeScript
21jevtypescript - Benchmarking Jev (Typesafe.ai) against a strong LLM on the Who&When Pro agent-failure-attribution benchmark (text subset).
TokenTrim/jev-agent-failure-benchmark · Python
20jevpython - A playground for TypeSafeAI's Jev Model
DeepBlueDynamics/typesafe-arena · Rust
20jevrust - 407jev-freeformAn observable raw-character chat experiment powered entirely by TypeSafe Jev Choice
kesku/jev-freeform · JavaScript
20jevjavascript - Reproducible Jev Ultrafast research-browser eval harness + field note (QC’d cases, suite runner, report generator). Not investment advice.
jgridifier/jev-research-eval · HTML
20jevhtml - Benchmarking TypeSafe's Jev decision model as a cost-efficient LLM router on RouterArena
TokenTrim/jev-routing-experiment · Python
22jevpython - Measures how well TypeSafe's RLCD-Jev model spots real secret credentials in file snippets
teyhouse/jev-secret-detection · Python
20jevpython - Zero-shot spam filtering with TypeSafe Jev Noul questions, compared with TF-IDF baselines
bitnovus/jev-spam-eval · Jupyter Notebook
20jevjupyter notebook - 423Jev-VisionOpen-weight step verifier for computer-use agents: calibrated ground/skip/effect/done judgments from screenshots in ~160 ms, plus a benchmark with environment-derived labels
sseanliu/Jev-Vision · Python
20jevpython - 615RISC-jeVI tortured Jev into being a RISC-V CPU.
i2cjak/RISC-jeV · Python
10jevpython - 456ask-jevUtilizing Jev, the RLCD-type model provided by TypeSafe AI, to independently and cheaply judge agentic coding sessions.
omni-/ask-jev · PowerShell
10jevpowershell - 482got-jevJev (TypeSafe AI) PoC through Game of Thrones
phureewat29/got-jev · TypeScript
10jevtypescript - 483gpt-vs-jevCompare GPT generated language with JEV structured Noul decisions on the same input.
TanayPadar/gpt-vs-jev · TypeScript
10jevtypescript - A small Next.js app for experimenting with TypeSafe AI's Jev model (System One)
Little-Planet-Labs/jev-playground · TypeScript
10jevtypescript - Jev (TypeSafe System One) × ASReview SYNERGY abstract screening demo — Choice/Noul vs gold labels
PistachioAIHQ/jev-synergy-screening · Python
11jevpython - AI benchmark on Japan's 2026 Common Test: Jev vs luna-none vs luna-low (static dashboard)
shibadogcap/kyotsu-ai-bench · HTML
10jevhtml - Typed-decision benchmark from PadFlow (land development SaaS): schemas, anonymized labeled rows, and a runner for confidence-calibrated models like TypeSafe Jev.
zsavage8/padflow-jev-evals · Python
10jevpython - Can a System One model steer music? Jev picks the plan (enums only); code renders sheet, audio and MIDI.
wustep/jev-playground · TypeScript
11jevtypescript - Jev research manuscript, evidence, and reproducible paper package
CompleteDotTech/paper-package · Python
10jev - 503jev-benchDoes the cited source actually say it? A 42-claim benchmark: Jev (TypeSafe System One) against GPT-5.4, Claude Sonnet 5 and Gemini 3.1 Pro.
TheWayWithin/jev-bench · Python
10jevpythonsystem-one - Next.js UI showing off TypeSafe's System One model (Jev) — parallel Noul judgments and a Choice-based citation checker, deployable to Vercel
Ashadeepa/typesafe-showcase · TypeScript
10jevtypescript - 646what-is-jevIndependent, source-linked research on TypeSafe AI's Jev (System One), with 947 rubric-scored public repositories, recurring patterns, datasets, and bilingual documentation.
g0runmezadam/what-is-jev · Python
10jevai-agentsbenchmarkbilingualclassificationdatasetdecision-modelllm - 585Jevs-GarageA garage full of tiny experiments for building critical systems with System One & Jev 🔧🧠⚡
JGalego/Jevs-Garage · Python
11jevai-demosai-safetydecision-intelligencedeveloper-toolsexplainable-aihuman-in-the-loopincident-response - Jev vs GPT-4.1 as synthetic survey respondents on Twin-2K-500. How you ask mattered more than which model you used.
jjd-lab/jev-synthetic-survey · Python
10jevbehavioral-economicsbenchmarkcalibrationdigital-twindigital-twinsllmllm-evaluation - 568jev-writerFind out which qualities of your writing actually predict engagement. Rates every post you have published against a pre-registered rubric using Jev's calibrated judgments, then tests those ratings against your real engagement numbers. Refuses to report findings your sample cannot support.
Kaos599/jev-writer · JavaScript
10jevagent-skillsai-gatewaycalibrated-probabilitiesclaude-codeclaude-skillcontent-analyticscreator-tools - Battleship against Jev, a model that answers in probabilities instead of text. Web game plus a CLI arena that plays it against general-purpose LLMs on identical fleets.
sah1l/jev-battleship · JavaScript
10jevbattleshipbenchmarkllmprompt-engineeringtypesafe-aiverceljavascript - 573jevcheckBehavioral contracts for TypeSafe Jev — pin production expectations, eval model upgrades, catch flips and confidence regressions.
sathariels/jevcheck · Python
10jevpython - 480ghost-userghost-user: TypeSafe Jev ecosystem repository.
shauryajain07/ghost-user · TypeScript
10jevtypescript - Jev Bayes, No? Testing TypeSafe AI's Jev against Bayesian-optimal strategies, and testing if Jev can effectivly use Bayesian priors.
TomRichner/can-jev-bayes · Python
10jevpython - typesafe-jev-traffic-demo: TypeSafe Jev ecosystem repository.
trycatchkamal/typesafe-jev-traffic-demo · Python
10jevpython - 564jev-tradejev-trade: TypeSafe Jev ecosystem repository.
Waxmell114514/jev-trade · Python
10jevpython - 581jevllmJev as an LLM (cz why not)
wisalkhanmv/jevllm · Python
10jevpython - 518jev-evalBenchmark TypeSafe Jev against any OpenRouter model on your own data.
4esv/jev-eval · Python
10jevpython - 517jev-driveJev + autonomous driving: structured decisions, multimodal baselines, recovery research, and measured API diagnostics.
Alpha-Harper-Franklin/jev-drive · Python
10jevautonomous-drivingmultimodalreplanningresearchvision-language-modelpython - 548jev-sandboxTest bench for TypeSafe's Jev
amr05008/jev-sandbox · TypeScript
10jevtypescript - 553jev-simJev-compatible /v1/systemone server reading typed decisions from LLM logits, benchmarked against TypeSafe's Jev on the same items via JevBench
dashbi1/jev-sim · Python
10jevbenchmarkcalibrationclassificationllmlogprobstypesafepython - 462BizzJevExperiments with TypeSafe/Jev semantic gates and a Semantic Operations Lab demo.
havietkok-sys/BizzJev · C#
10jevc# - automatically optimizing the instructions and decision criteria of TypeSafe Jev Choice from labeled data
j341nono/jev-prompt-optimization · Python
10jevpython - 523jev-graphragSmall demos + use-case backlog: TypeSafe AI's Jev as a calibrated decision layer for GraphRAG pipelines on Neo4j.
neo4j-field/jev-graphrag · Python
10jevpython - 516jev-doomWatch Jev play Freedoom in a local dashboard. TypeSafe direct and Vercel AI Gateway, inspectable decisions, and bounded spending.
olivier-motium/jev-doom · Python
10jevai-agentsdoomfreedoomvercel-ai-gatewaypython - A WebGL demo where you play the card game Speed against a CPU whose brain is TypeSafe AI's Jev. The whole point of the app is to measure and show Jev's decision speed and decision accuracy in real time.
tubone24/jev-practice-speed · JavaScript
10jevcard-gamejev-aijavascript - Jev (TypeSafe System One) decision tools + live verification benchmark for DeepSeek Harness: jev_decision (choice/score/noul) and jev_verify, honest by design.
xienda/dsh-jev-verify · JavaScript
10jevjavascript - jev-music-theory-1: TypeSafe Jev ecosystem repository.
adammichaelwood/jev-music-theory-1 · TypeScript
10jevtypescript - 562jev-testPre-registered benchmark: can a 2B local model (Gemma 4 E2B) answer web questions without making things up when a decision model (TypeSafe Jev) makes every call? SearXNG for search, MemPalace for verbatim memory, seven arms including open local judges. Spec and thresholds fixed before any run.
clduab11/jev-test · Python
10jevbenchmarkgemmahallucinationmempalaceragsearxngsmall-language-models - Measure what Jev can actually do before you build on it. Graded findings, ruled-out candidates, and recipes with stop-conditions. 0 promotions — on purpose.
JYeswak/jev_playground · Python
10jevagentsbenchmarksevaluationllmtypesafepython - Using Jev to test how well it predicts financial markets(just like most llms as of september 2026, it doesnt do that good)
thodoh1/FinancialPredictionJev · Python
00jevpython - An evaluation of typesafe AI chess. As it turns out, the AI isn't doing really well even though chess is not a particularly open-ended game. Still, it's only a prototype and this probably wasn't optimzied for games.
AliceRoselia/Typesafe_chess_eval · Python
00jevpython - Experimental design harness: a small decision model (Jev) picks the design in about a second, a traditional LLM (Luna) only writes the words. With and without it.
LamplighterPaul/forma-system1-experiment · TypeScript
00jev - 718ground-zeroDecision library to detect and classify AI hallucinations, powered by Jev AI.
zavocc/ground-zero · Python
00jev - Does Jev predict stock returns from news? It reads the news well; there is no tradeable alpha. Three arms separate reading from recall.
Gaurav-Gosain/jev-alpha-bench · Go
00jevgo - Jev (TypeSafe) vs. Gemini 3.8 Flash vs. GPT-5.6 Luna na anotação estruturada de sentenças do TJSP: qualidade, tempo e custo
lab-dados/jev-anotacao-sentencas · Python
00jevpython - Challenge Atari with Jev: structured decisions, value questions, and replayable experiments
memorysaver/jev-atari-lab · Python
00jevdemo - 762jev-benchTypeSafe / Jev community project: thomasschafer/jev-bench.
thomasschafer/jev-bench · Python
00jev - Position paper: the Hidden-Markov and fuzzy primitives missing from TypeSafe AI's Jev and System-One decision models. Two lemmas, one principle (Deferred Crispification), one architecture (BSF-S1).
dnakhoa/jev-deferred-crispification · TeX
00jevtexcalibrationfuzzy-logichidden-markov-model - 804jev-demosDemos to test the effectiveness of TypeSafe's "Jev" System One Model
Bud-ro/jev-demos · Dart
00jevdart - 806jev-dev同じ発言を jev と LLM の両方に判定させ、感情の変動値のズレと応答速度を1画面で見比べるデモ(affectus + Vercel AI Gateway)
n-yokomachi/jev-dev · TypeScript
00jevtypescript - typesafe.ai model jev finance benchmark
hifizz/jev-finance-benchmark
00jev - Can Jev pick the winner of a real headline A/B test? 64.5% across 10,984 Upworthy randomized experiments, 74.7% when the difference was decisive.
Gaurav-Gosain/jev-headline-bench · Go
00jevgo - Jev (TypeSafe) 性能評価プロジェクト — 日本郵便 KEN_ALL をマスタに、AI SDK 経由の Jev が住所のあいまい一致にどこまで使えるかを検証
smasato/jev-jp-address · TypeScript
00jevtypescript - 844jev-labTypeScript experiments, evaluations, and latency benchmarks for TypeSafe's Jev model
Menny1337/jev-lab · TypeScript
00jevtypescript - A small reproducible MuJoCo pilot comparing Jev, Claude Haiku, and reactive rules for pick-and-place.
tryaksh/jev-pick-and-place-study · Python
00jevpython - 896jev-report发明 RLHF 的人,这次做了个不会说话的模型:Jev 独立研究报告。52 页 PDF + 50 条中文实测复现包 + 143 条可回溯数据表
HackSing/jev-report · Python
00jevpythonai-researchchinesellm-evaluation - A small second eval for shadcn-ui/lint that uses TypeSafe's Jev to judge the linter's own output.
blas0/jev-shadcn-lint-eval · JavaScript
00jevjavascript - Application of TypeSafe Jev (noul judgment primitive) on the collusion.wiki corpus: agent vs human page authorship, head-to-head vs local Qwen3.8-Flash-Next
sypherin/jev-trace-classifier · Python
00jevpython - 954jev-vs-lunaReproducible Jev vs Luna review-classification benchmark with measured accuracy, latency, and costs.
mameli/jev-vs-luna · Python
00jev - 1067roverlabA 3D planetary rover sandbox for experimenting with autonomous decisions using TypeSafe AI.
juancamiloqhz/roverlab · TypeScript
00jevtypescript - Evaluating TypeSafe's Jev as a fast monitor and action gate for agent sabotage in SHADE-Arena, compared with Gemini 2.5 Flash/Pro.
nican2018/shade-arena-jev-monitor · Python
00jevpython - 1098terrariumA sandbox where a TypeSafe System One model presses the controls of a small creature. Code runs the world.
TheGali/terrarium · JavaScript
00jevjavascript - Charts: TypeSafe Jev evaluated on Thai standardized exams vs 110 other models
vehas/thaiexam-jev-charts · HTML
00jevhtml - 1105tiny-jevTypeSafe / Jev community project: karimatayuta/tiny-jev.
karimatayuta/tiny-jev · HTML
00jev - 1108trading-bot-jevCrypto trading bot on Binance testnet using TypeSafe (Jev) to judge news
Spykoninho/trading-bot-jev · TypeScript
00jev - 1139typesafe-oraclesEvaluating TypeSafe's System One primitives (Choice/Score/Noul) — where a typed oracle beats an LLM call
trophee-bot/typesafe-oracles · JavaScript
00jevjavascript - 891jev-projectsSmall demos of Jev (TypeSafe) through the Vercel AI Gateway: wiki race, town of agents, bullet chess, and more
az9713/jev-projects · JavaScript
00jevjavascriptdemo - 803jev-demoHistorical paper-trading simulator for evaluating TypeSafe AI JEV decisions
co1smos/jev-demo · Python
00jevpythondemo - 740jev_practiceJev (TypeSafe System One) と LLM に同じゲームを打たせて、レイテンシ・コスト・判断の質を比べる練習台
ryuchan00/jev_practice · Python
00jevpythonsystem-one - Does TypeSafe's Jev keep its accuracy and calibration on Russian? Independent RU vs EN audit (ECE, reliability diagrams, paired bootstrap) on parallel human-labelled data.
AHTOOOXA/jev-cyrillic-audit · Python
00jevpython - jev-llm-benchmark: TypeSafe Jev ecosystem repository.
Chronona/jev-llm-benchmark · TypeScript
00jevtypescript - Does a System One model actually beat keyword matching? A reproducible benchmark on catching disguised duplicate thesis titles. 24 cases, real production baseline, raw data and charts included.
devnolife/jev-vs-tfidf-benchmark · TypeScript
00jevbenchmarkindonesiainformation-retrievalllmplagiarism-detectionstructured-outputtf-idf - 772jev-btzscjev-btzsc: TypeSafe Jev ecosystem repository.
Gazer2020/jev-btzsc · Python
00jevpython - Evaluate typed AI decisions on labeled French-language cases.
gbesse/jev-banc-francais · JavaScript
00jevcalibrationevalsfrancetypesafe-aijavascript - 782jev-codebookQualitative coding at scale with Jev: apply a codebook to open-ended text, review uncertain items, measure agreement against your human coders.
gbesse/jev-codebook · Python
00jevhuman-in-the-loopinter-rater-reliabilityopen-sourcepythonqualitative-researchtext-classification - 788jev-crowdsimRun a declared factorial audience grid through typed Jev reactions and expose disagreement.
gbesse/jev-crowdsim · JavaScript
00jevaudience-researchfactorial-designnodejsopen-sourcesimulationsynthetic-datajavascript - 902jev-roastScore declared writing dimensions and cite only exact source spans for weak results.
gbesse/jev-roast · JavaScript
00jevcontent-qualityexplainable-ainodejsopen-sourcetext-evaluationwriting-analysisjavascript - 904jev-screenTitle and abstract screening for systematic reviews with Jev: explicit criteria, include/exclude/maybe with reasons, PRISMA counts, RIS export, recall against human screeners.
gbesse/jev-screen · Python
00jevevidence-synthesishuman-in-the-loopliterature-screeningopen-sourceprismapythonsystematic-review - 866jev-nflBefore every NFL snap, a decision-only AI model (TypeSafe Jev) calls run or pass and go/punt/kick on fourth down, graded live against the coach.
gregjonesio/jev-nfl · JavaScript
00jevjavascript - 930jev-testbedJev (TypeSafe System One) 테스트베드 — 클라우드 API와 로컬 셀프호스팅(jeff/GLiFormer) 양쪽 실행 예제 및 실측 결과
hulryung/jev-testbed · Python
00jevpython - Local JEV experiments: route decisions, Minesweeper solvers, and drone simulation
imom39a/jev-playground · Python
00jevpython - Fifty real-world financial use cases for TypeSafe's Jev model: typed, structured LLM answers over ledgers, fraud, portfolios, trades and filings, each graded against data where the right answer is known.
IslamBaraka90/jev-typesafe-real-financial-use-cases · JavaScript
00jevai-agentsbacktestingfinancial-datafintechjavascriptllmllm-evaluation - A live, graphical dojo for TypeSafe's Jev (System One) typed decision model — routing, a Tetris-playing agent, parallel swarms, and an honest Jev-vs-Claude gauntlet.
lafollett-labs/typesafe-jev-dojo · TypeScript
01jevai-agentscanvasdecision-modelopenroutersystem-onetetristypesafe - 674calibrantDoes your model's confidence mean anything on your data? Calibration layer for typed probabilistic decisions — reliability, ECE, Brier, recalibration maps and cost-aware thresholds from decisions + outcomes.
Maher-Reven/calibrant · TypeScript
00jevtypescript - 744jev-acento¿Jev entiende tu acento? Pre-registered audit of TypeSafe AI's Jev on Spanish — accuracy, calibration and token cost — plus a CLI to run the same comparison on your own labelled data.
marcosmartinez/jev-acento · Python
01jevbenchmarkcalibrationexpected-calibration-errorllm-evaluationnlppre-registrationreproducible-research - 828jev-hankoMeasuring TypeSafe AI's Jev on 41-clause contract review (CUAD, 20,500 decisions) against fast, cheap LLMs — latency, cost and F1
matu79go/jev-hanko · Python
00jevbenchmarkcontract-reviewcuadlatencylegal-aillmopenrouter - Small experiments with Jev by TypeSafe
nak1b/jev-experiments · TypeScript
00jevaiai-experimentsjev-aitypesafe-aitypescript - demo trend seracher using jev
nhchoi98/demo_trend_searcher · TypeScript
00jevtypescript - 801jev-demoIndependent demo of TypeSafe's Jev model: typed decisions with probabilities, measured side by side with OpenAI on support-ticket triage. Live local app plus a recorded replay page.
ogamircs/jev-demo · HTML
00jevdemollm-benchmarkopenaistructured-outputstypesafe-aihtml - 784jev-computerAn 8-bit computer built from one yes/no question asked to Jev (TypeSafe) — 24,511 NAND gates from a single API call
RiwRiwara/jev-computer · Python
00jevpython - Independent playground for TypeSafe AI Jev decision models: typed decisions, support-ticket routing, reproducible evaluations, and a local browser demo.
STiFLeR7/Jev-LLM-Playground · JavaScript
00jevai-evaluationbenchmarkingdecision-modelsjavascriptnodejsplaygroundstructured-decisions - 919jev-snap-labTiny inputs. Instant decisions. A small experimental playground for exploring fast, probabilistic decisions with Jev.
takafumikobayashi/jev-snap-lab · TypeScript
00jevtypescript - Local Jev evaluation workbench: datasets, typed questions, threshold simulation and run comparison
Tomdachs/jev-replay-lab · TypeScript
00jevtypescript - english-2-sql: TypeSafe Jev ecosystem repository.
trivektor/english-2-sql · JavaScript
00jevjavascript - 960jevaljeval: open-source evaluations for AI outputs and agents, judged by Jev
vrash/jeval · TypeScript
00jevtypescript - 1031open-system-oneAn independent, reproducible benchmark of TypeSafe's Jev against open, CPU-only alternatives — 10,000 decisions, all raw results published.
zhlei07/open-system-one · Python
01jevbenchmarkcpu-inferencecross-encoderembeddingstext-classificationtypesafezero-shot-classification - 656ask-twiceask-twice: TypeSafe Jev ecosystem repository.
aarongunasingh/ask-twice · Python
00jevpython - A learning scaffold for TypeSafe AI's System One models: eval harness plus a measured, plain-language comparison of the Jev decision model vs an LLM stand-in on 24 real operational decisions. All numbers reproducible from committed run files.
andreaserradev-gbj/jev-access-day · TypeScript
00jevaibenchmarkdecision-supportevaluationllmtype-safetytypesafe-ai - 800jev-demojev-demo: TypeSafe Jev ecosystem repository.
aoprisan/jev-demo · Rust
00jevrust - 756jev-arcadeCan a System One model play arcade games? TypeSafe's Jev plays Tetris, Snake and 2048 — benchmarked against random and heuristic baselines.
CankatSarac/jev-arcade · Python
00jevpython - 928jev-testPlayground for TypeSafe's Jev System One model (Next.js)
Dillettant/jev-test · TypeScript
00jevtypescript - 710fraud-jevfraud-jev: TypeSafe Jev ecosystem repository.
frankied003/fraud-jev · TypeScript
00jevtypescript - A 60-game benchmark of TypeSafe's Jev evaluation model playing Battleship. The model matches plain code; it does not beat it.
ickas/battleship-vs-jev · TypeScript
00jevaibattleshipbenchmarkevaluation-metricsllm-evaluationtypescript - A cached Jev prior for active learning: reusable rank fusion, matched ASReview controls, and a no-key evidence replay.
joaovaleri/jev-shortlist · Python
00jevactive-learningasreviewpythonreciprocal-rank-fusionreproducible-researchsystematic-reviewtypesafe - 987JevScopeJevScope: TypeSafe Jev ecosystem repository.
KaushikKC/JevScope · TypeScript
00jevtypescript - Jev × obniz LED: Physical AI Hello World — text → typed decisions (TypeSafe System One) → WS2812B LEDs
kofujimura/jev-obniz-led · TypeScript
00jevtypescript - 977jevmazejevmaze: TypeSafe Jev ecosystem repository.
kt3k/jevmaze · TypeScript
00jevtypescript - 842jev-labReal browser-agent safety evaluation: Jev versus a baseline on benign and injected tasks
mjyoke1111/jev-lab · TypeScript
00jevtypescript - Experimental PoC for Jev translation checking: bilingual benchmarks, prompt comparisons, re-verification, and MAGI voting.
mshk/jev-translation-checker · JavaScript
00jevjavascript - A pre-registered field trial of Jev (TypeSafe's judgment model) on a second brain and Claude Code history: 20 tests, bars written first, failures included, and the tools to repeat it.
NaluKicks-808/jev-field-trial · Python
00jevai-agentsclaude-codeevaluationsecond-braintypesafepython - 867jev-no-enemReproducible benchmark evaluating TypeSafe AI's Jev (System One paradigm) on Brazil's ENEM 2025 standardized exam. Evaluates typed decision-making, domain-specific accuracy, and RLCD uncertainty calibration against open LLM baselines with an interactive GitHub Pages dashboard.
patryckalves/jev-no-enem · Python
00jevbenchmarkbrasilenemllmrlcdsystem-one-modelspython - jev-integration-report: TypeSafe Jev ecosystem repository.
piratchai/jev-integration-report
00jev - Reproducible Tetris decision benchmark comparing TypeSafe Jev with Claude Haiku
planstack-ai/jev-tetris-benchmark · TypeScript
00jevai-evaluationnextjstetristypesafe-aivercel-ai-gatewaytypescript - 1168WHAT-s-Up-jevWHAT-s-Up-jev: TypeSafe Jev ecosystem repository.
Pragyan330/WHAT-s-Up-jev · Python
00jevpython - 998jevyJev-style typed-decision model distilled from official Jev. 118M, EN+CN, trains on a 4GB GPU in 10 minutes.
slatinwine/jevy · Python
00jevpython - jev-noul-vs-choice: TypeSafe Jev ecosystem repository.
TakumiNoguchi2004/jev-noul-vs-choice · Python
00jevpython - 889jev-pocTypeSafe AI の判定モデル jev に 2048 を遊ばせる PoC(Go CLI + Cloudflare Workers の Web デモ)
tatsuo48/jev-poc · Go
00jevgo - 997jevxJev (TypeSafe System One) research: API notes, benchmarks, community experiments, agent-loop patterns
umgbhalla/jevx · Python
00jevpython - 732jevIndependent research notes toward an open Jev-like decision model: public facts, API contract, training and eval plan.
WiredMind2/jev · Python
00jevpython - Experiments with TypeSafe's Jev System One model
xavierforge/jev_experiments · JavaScript
00jevjavascript - Live demo showing why loop-speed classification matters: Jev vs LLMs on the same events, same clock, honest scorecard.
yshraj/jev-traffic-race · TypeScript
00jevtypescript - Label every sentence of a document with calibrated probabilities from Jev, rendered as a heatmap
abhishekmishragithub/semantic-microscope · Python
00jevllmpythontypesafevisualization - 961JEValuateAuto-marking maths scripts with Jev (TypeSafe System One): 2,054 scripts, 96.6% agreement with human markers
Akeel-Majeed/JEValuate · TypeScript
00jevtypescript - security-sandbox-jev: TypeSafe Jev ecosystem repository.
altanapps/security-sandbox-jev · Python
00jevpython - 1020MiraveMirave: TypeSafe Jev ecosystem repository.
edoigtrd/Mirave · Python
00jevpython - 1060reflex-jevTraining demonstration of Jev in a dispatch services command center scenario.
fullcolorcoder/reflex-jev · TypeScript
00jevtypescript - Empirical experiments and API research for TypeSafe's Jev System One model trained using RLCD
BipinRajC/Jev-api-experiments · Python
00jevapiexperimentalrlcdtypesafe-aipython - jev-experiments: TypeSafe Jev ecosystem repository.
harlanljones/jev-experiments · JavaScript
00jevjavascript - jev-experiments: TypeSafe Jev ecosystem repository.
pavan142/jev-experiments · TypeScript
00jevtypescript