TypeSafe Jev / System One 的免费 AI 精选目录按用途分类的 GitHub 开源项目

数据更新

  • Community Jev-like model and tooling experiment inspired by TypeSafe System One.

    vinnylarouge/jevlike · Python

    1,283113
    jevdemomodel
  • jev-experiments: TypeSafe Jev ecosystem repository.

    dabit3/jev-experiments · TypeScript

    38128
    jevtypescript
  • Camera-only autonomous drone in MuJoCo with a small judgment model (TypeSafe Jev) in the loop at 2.5Hz

    RomanSlack/jev-drone · Python

    17624
    jevdemo
  • JevBench v1 - a benchmark for Jev-class typed decision models: smart, cheap, fast, reliable, open.

    fstandhartinger/jevbench · Python

    11812
    jevpython
  • Using Jev as an evaluator.

    danielgshea/jev-as-a-judge · Python

    8514
    jevpython
  • Agent evals and guardrails as Jev decisions: one request per trace, a fraction of a cent, fast enough for the agent loop. Runs locally with Kev or Laya.

    openlayer-ai/jevals · Python

    835
    jevagentsevalsguardrailsllmllm-evaluationragastypesafe
  • JEV-inspired parallel decisions for CUDA LLMs. One context, many decisions. vLLM API, game-agent examples, and reproducible benchmarks.

    kikoncuo/jevfire · JavaScript

    645
    jevjavascriptcudagame-aiinference
  • Open reproduction of TypeSafe Jev: a 150M typed decision engine (noul/choice/score in one non-autoregressive pass, calibrated confidence). 0.697 vs Jev's 0.727, 2.5x better calibrated, 4x faster, free. Trains on a Colab T4 in 30 min.

    intikhab49/open-jev-typed-decision-engine · Python

    432
    jevagent-observabilitycalibrationcolabdecision-modelexpected-calibration-errorhuggingfacellm-alternative
  • The highly anticipated open-source repository for JEV as Policy enables one-click setup of the simulation environment. Evaluations of Astra + JEV on benchmarks such as RoboTwin will also be released soon.

    YuanKJing/Jev-as-Policy · Python

    402
    jevpython
  • Typed JSON inference with DiffusionGemma, with Every and Jev benchmark results

    JoshuaSP/open-jev · Python

    392
    jevpython
  • This is a LLM Gateway that mimics typesafe ai structured output. Like an imposter Jev.

    iammrduncan/typesafe-ai-benchmark · TypeScript

    385
    jev
  • Open replica of TypeSafe's Jev: typed calibrated decisions in one forward pass, on Gemma 4 E2B / Gemma 3 270M (Modal)

    mithalouni/system-one-open · Python

    354
    jevpython
  • Local bilingual probability decisions from context, questions, and candidate answers. Independent research preview inspired by TypeSafe Jev.

    zhihz/openjev · Python

    333
    jevpython
  • Independent, evidence-based map of when TypeSafe's Jev actually holds up vs. breaks down — real API-call receipts, not a leaderboard. 中文為主的雙語 repo。

    Zaious/jev-capability-atlas · Python

    266
    jevai-agentsbenchmarkcalibrationllm-evaluationmachine-learningtypesafezh-tw
  • Unofficial study: Jev-style parallel typed decisions on stock 1.5B-8B models on an Apple Silicon laptop. Benchmarks, research notes, and a Hugging Face Space demo.

    rorshopping/jev-on-a-laptop · Python

    241
    jevpython
  • Jev vs Gemini 3.8 Flash: labelling 1,000 app reviews, 4.1× faster and 7× cheaper

    goodrahstar/jev-column-race · JavaScript

    233
    jevjavascript
  • Probability-aware evaluation for typed decision models: calibration, selective risk, latency, and reproducible benchmarks.

    AbdelStark/jev-benchmarks · Python

    173
    jevpythonbenchmarkingcalibrationevaluation
  • Measures what your Jev classifier's confidence is really worth, and sets the human hand-off line from what a mistake costs.

    rlaope/jeval · Python

    171
    jevpython
  • A stronger one-pass scorer over a variable list of text options. Hashed n-gram encoder, rival-aware attention, gated head, temperature scaling — with a head-to-head benchmark vs the jevlike starter design.

    olanotolu/jevbetter · Python

    143
    jevpython
  • An experimental JEV-powered framework for forecasting short-term stock price direction from structured market data.

    sosopop/jev_stock · Python

    133
    jevpythonfinancial-datafinancial-machine-learningllm
  • Interactive experiments with TypeSafe Jev, from support routing to 3D driving simulations with real AI decisions and visible sensor inputs.

    kavehmz/typesafe-playground · JavaScript

    134
    jevjavascript
  • Backtest Jev (TypeSafe) as a BUY/SELL/HOLD trader on NQ L10 order-book data

    justinhe16/trade-jev · Python

    83
    jevpython
  • Reproducible calibration and selective-risk benchmarks for Jev/TypeSafe decisions in DSPy workflows

    jmanhype/jev-dspy-lab · Python

    80
    jevpython
  • Can a decision model beat dedicated rerankers? TypeSafe Jev vs Cohere Rerank 4 vs ZeroEntropy zerank-2 vs a chat-model baseline: 14 datasets, every raw API response, bootstrap ranges on every gap.

    anessbelbati/jev-rerank-bench · Python

    70
    jevpython
  • Connect JEV to MCP clients and compare its judgments against general-purpose LLMs using shared datasets and measurable accuracy.

    arunav25/jev-mcp · JavaScript

    70
    jevjavascript
  • ui-generator-instinct-jev: TypeSafe Jev ecosystem repository.

    joevidev/ui-generator-instinct-jev · TypeScript

    70
    jevtypescript
  • Reproducible early-access evaluation of Jev on Korean understanding and medical text, with runtime and cost evidence

    mahlernim/jev-korean-benchmark · Python

    60
    jevpython
  • A word-level language model whose output layer is Jev: n-gram drafter, Noul chunk verification, bits-per-token eval

    y0usaf/jev-lm · TypeScript

    61
    jevtypescript
  • Benchmarks and a playground for TypeSafe's Jev (System One) model: chess, and who-is-the-player-talking-to for speech-to-text game NPCs

    wondertwins/jev-benchmark · Python

    61
    jevpython
  • High-speed recursive AI Elo tournament engine powered by Jev and Swiss matchmaking

    opaielsheikh/ai-elo-ranker · Python

    52
    jevpython
  • A show-and-tell capability study for Jev, TypeSafe's System One decision model.

    lbotinelly/jev-little-airways · HTML

    50
    jevhtml
  • Jev (TypeSafe) vs Claude Haiku 4.5 on 2 000 phishing emails: accuracy, calibration, latency, cost. Reproducible benchmark.

    anisselbd/jev-phishing-bench · Python

    50
    jevpython
  • Jev-shaped typed-decision model (state + Choice/Score/Noul questions -> calibrated probabilities, one pass) on ModernBERT / DeBERTa / LLaDA-MoE, with measured latency, accuracy, calibration and training cost

    kotoba-lang/typed-decisions · Python

    50
    jev
  • An instrumented 2048 web lab where every move is a Jev (TypeSafe AI System One) Choice, with no heuristic fallback | 用 Jev 决策模型驱动每一步的 2048 网页实验台,概率、置信度、延迟与成本全部摊开可见,且刻意不做启发式兜底

    ARCJ137442/jev-2048 · TypeScript

    51
    jev2048llm-evaluationsystem-onetyped-decisionstypesafe-aitypescript
  • Jev-style calibrated decision model (Choice/Score/Noul) on Qwen3.5-0.8B

    shamazharikh/qwen-rlcd · Python

    41
    jevpython
  • Challenge the Jev's intelligence in Rubik Cube puzzles

    0xtrou/rubikjev · TypeScript

    40
    jevtypescript
  • Independent Jev 1.13.0 behavior study: report, controlled prompt experiments, raw results, and offline verification.

    RINNECODER/jev-behavior-study · Python

    30
    jevpython
  • Blind security benchmarks for Jev, TypeSafe's System One model: prompt injection and vulnerable code detection, built on jev-go

    Gaurav-Gosain/jev-sec-bench · Go

    30
    jevgo
  • A playground for experiments around Jev, TypeSafe's System One model.

    markjaquith/typesafe-ai-playground · Rust

    31
    jevrust
  • Jev (TypeSafe) exploratory thread: claim audit, live demos, and runnable code

    SamuelSacco/jev-exploration · Python

    30
    jevpython
  • Three measured experiments on RAG hallucination: quote-checking, TypeSafe's Jev, and IBM's STAIR. 850+ graded questions, raw responses included.

    aryanchauhanoffical/no-hallucination · Python

    30
    jevbenchmarkevaluationhallucinationllmragretrievaltypesafe-ai
  • About calibrating Jev for code reviews

    Selmar/typesafe-jev-calibrate-for-code-review · Python

    30
    jevpython
  • JEV Reinforcement Learning: four classic games trained with JEV-powered rewards, reproducible experiments and checkpoint replays.

    Bring-AI/jev-rl · Python

    30
    jevpython
  • jev-plays: TypeSafe Jev ecosystem repository.

    mansicer/jev-plays · Python

    30
    jevpython
  • Measure when to use Jev and other models on your data, then route accordingly.

    FirasSX914/calibre · Python

    20
    jevpythonbenchmarkcalibrationconfidence
  • Experimental multi-horizon BTC signal generator using TypeSafe Jev probabilities and Binance market data.

    WebGrga/btc-jev-signal · TypeScript

    21
    jevtypescript
  • Benchmarking Jev (Typesafe.ai) against a strong LLM on the Who&When Pro agent-failure-attribution benchmark (text subset).

    TokenTrim/jev-agent-failure-benchmark · Python

    20
    jevpython
  • A playground for TypeSafeAI's Jev Model

    DeepBlueDynamics/typesafe-arena · Rust

    20
    jevrust
  • An observable raw-character chat experiment powered entirely by TypeSafe Jev Choice

    kesku/jev-freeform · JavaScript

    20
    jevjavascript
  • Reproducible Jev Ultrafast research-browser eval harness + field note (QC’d cases, suite runner, report generator). Not investment advice.

    jgridifier/jev-research-eval · HTML

    20
    jevhtml
  • Benchmarking TypeSafe's Jev decision model as a cost-efficient LLM router on RouterArena

    TokenTrim/jev-routing-experiment · Python

    22
    jevpython
  • Measures how well TypeSafe's RLCD-Jev model spots real secret credentials in file snippets

    teyhouse/jev-secret-detection · Python

    20
    jevpython
  • Zero-shot spam filtering with TypeSafe Jev Noul questions, compared with TF-IDF baselines

    bitnovus/jev-spam-eval · Jupyter Notebook

    20
    jevjupyter notebook
  • Open-weight step verifier for computer-use agents: calibrated ground/skip/effect/done judgments from screenshots in ~160 ms, plus a benchmark with environment-derived labels

    sseanliu/Jev-Vision · Python

    20
    jevpython
  • I tortured Jev into being a RISC-V CPU.

    i2cjak/RISC-jeV · Python

    10
    jevpython
  • Utilizing Jev, the RLCD-type model provided by TypeSafe AI, to independently and cheaply judge agentic coding sessions.

    omni-/ask-jev · PowerShell

    10
    jevpowershell
  • Jev (TypeSafe AI) PoC through Game of Thrones

    phureewat29/got-jev · TypeScript

    10
    jevtypescript
  • Compare GPT generated language with JEV structured Noul decisions on the same input.

    TanayPadar/gpt-vs-jev · TypeScript

    10
    jevtypescript
  • A small Next.js app for experimenting with TypeSafe AI's Jev model (System One)

    Little-Planet-Labs/jev-playground · TypeScript

    10
    jevtypescript
  • Jev (TypeSafe System One) × ASReview SYNERGY abstract screening demo — Choice/Noul vs gold labels

    PistachioAIHQ/jev-synergy-screening · Python

    11
    jevpython
  • AI benchmark on Japan's 2026 Common Test: Jev vs luna-none vs luna-low (static dashboard)

    shibadogcap/kyotsu-ai-bench · HTML

    10
    jevhtml
  • Typed-decision benchmark from PadFlow (land development SaaS): schemas, anonymized labeled rows, and a runner for confidence-calibrated models like TypeSafe Jev.

    zsavage8/padflow-jev-evals · Python

    10
    jevpython
  • Can a System One model steer music? Jev picks the plan (enums only); code renders sheet, audio and MIDI.

    wustep/jev-playground · TypeScript

    11
    jevtypescript
  • Jev research manuscript, evidence, and reproducible paper package

    CompleteDotTech/paper-package · Python

    10
    jev
  • Does the cited source actually say it? A 42-claim benchmark: Jev (TypeSafe System One) against GPT-5.4, Claude Sonnet 5 and Gemini 3.1 Pro.

    TheWayWithin/jev-bench · Python

    10
    jevpythonsystem-one
  • Next.js UI showing off TypeSafe's System One model (Jev) — parallel Noul judgments and a Choice-based citation checker, deployable to Vercel

    Ashadeepa/typesafe-showcase · TypeScript

    10
    jevtypescript
  • Independent, source-linked research on TypeSafe AI's Jev (System One), with 947 rubric-scored public repositories, recurring patterns, datasets, and bilingual documentation.

    g0runmezadam/what-is-jev · Python

    10
    jevai-agentsbenchmarkbilingualclassificationdatasetdecision-modelllm
  • A garage full of tiny experiments for building critical systems with System One & Jev 🔧🧠⚡

    JGalego/Jevs-Garage · Python

    11
    jevai-demosai-safetydecision-intelligencedeveloper-toolsexplainable-aihuman-in-the-loopincident-response
  • Jev vs GPT-4.1 as synthetic survey respondents on Twin-2K-500. How you ask mattered more than which model you used.

    jjd-lab/jev-synthetic-survey · Python

    10
    jevbehavioral-economicsbenchmarkcalibrationdigital-twindigital-twinsllmllm-evaluation
  • Find out which qualities of your writing actually predict engagement. Rates every post you have published against a pre-registered rubric using Jev's calibrated judgments, then tests those ratings against your real engagement numbers. Refuses to report findings your sample cannot support.

    Kaos599/jev-writer · JavaScript

    10
    jevagent-skillsai-gatewaycalibrated-probabilitiesclaude-codeclaude-skillcontent-analyticscreator-tools
  • Battleship against Jev, a model that answers in probabilities instead of text. Web game plus a CLI arena that plays it against general-purpose LLMs on identical fleets.

    sah1l/jev-battleship · JavaScript

    10
    jevbattleshipbenchmarkllmprompt-engineeringtypesafe-aiverceljavascript
  • Behavioral contracts for TypeSafe Jev — pin production expectations, eval model upgrades, catch flips and confidence regressions.

    sathariels/jevcheck · Python

    10
    jevpython
  • ghost-user: TypeSafe Jev ecosystem repository.

    shauryajain07/ghost-user · TypeScript

    10
    jevtypescript
  • Jev Bayes, No? Testing TypeSafe AI's Jev against Bayesian-optimal strategies, and testing if Jev can effectivly use Bayesian priors.

    TomRichner/can-jev-bayes · Python

    10
    jevpython
  • typesafe-jev-traffic-demo: TypeSafe Jev ecosystem repository.

    trycatchkamal/typesafe-jev-traffic-demo · Python

    10
    jevpython
  • jev-trade: TypeSafe Jev ecosystem repository.

    Waxmell114514/jev-trade · Python

    10
    jevpython
  • Jev as an LLM (cz why not)

    wisalkhanmv/jevllm · Python

    10
    jevpython
  • Benchmark TypeSafe Jev against any OpenRouter model on your own data.

    4esv/jev-eval · Python

    10
    jevpython
  • Jev + autonomous driving: structured decisions, multimodal baselines, recovery research, and measured API diagnostics.

    Alpha-Harper-Franklin/jev-drive · Python

    10
    jevautonomous-drivingmultimodalreplanningresearchvision-language-modelpython
  • Test bench for TypeSafe's Jev

    amr05008/jev-sandbox · TypeScript

    10
    jevtypescript
  • Jev-compatible /v1/systemone server reading typed decisions from LLM logits, benchmarked against TypeSafe's Jev on the same items via JevBench

    dashbi1/jev-sim · Python

    10
    jevbenchmarkcalibrationclassificationllmlogprobstypesafepython
  • Experiments with TypeSafe/Jev semantic gates and a Semantic Operations Lab demo.

    havietkok-sys/BizzJev · C#

    10
    jevc#
  • automatically optimizing the instructions and decision criteria of TypeSafe Jev Choice from labeled data

    j341nono/jev-prompt-optimization · Python

    10
    jevpython
  • Small demos + use-case backlog: TypeSafe AI's Jev as a calibrated decision layer for GraphRAG pipelines on Neo4j.

    neo4j-field/jev-graphrag · Python

    10
    jevpython
  • Watch Jev play Freedoom in a local dashboard. TypeSafe direct and Vercel AI Gateway, inspectable decisions, and bounded spending.

    olivier-motium/jev-doom · Python

    10
    jevai-agentsdoomfreedoomvercel-ai-gatewaypython
  • A WebGL demo where you play the card game Speed against a CPU whose brain is TypeSafe AI's Jev. The whole point of the app is to measure and show Jev's decision speed and decision accuracy in real time.

    tubone24/jev-practice-speed · JavaScript

    10
    jevcard-gamejev-aijavascript
  • Jev (TypeSafe System One) decision tools + live verification benchmark for DeepSeek Harness: jev_decision (choice/score/noul) and jev_verify, honest by design.

    xienda/dsh-jev-verify · JavaScript

    10
    jevjavascript
  • jev-music-theory-1: TypeSafe Jev ecosystem repository.

    adammichaelwood/jev-music-theory-1 · TypeScript

    10
    jevtypescript
  • Pre-registered benchmark: can a 2B local model (Gemma 4 E2B) answer web questions without making things up when a decision model (TypeSafe Jev) makes every call? SearXNG for search, MemPalace for verbatim memory, seven arms including open local judges. Spec and thresholds fixed before any run.

    clduab11/jev-test · Python

    10
    jevbenchmarkgemmahallucinationmempalaceragsearxngsmall-language-models
  • Measure what Jev can actually do before you build on it. Graded findings, ruled-out candidates, and recipes with stop-conditions. 0 promotions — on purpose.

    JYeswak/jev_playground · Python

    10
    jevagentsbenchmarksevaluationllmtypesafepython
  • Using Jev to test how well it predicts financial markets(just like most llms as of september 2026, it doesnt do that good)

    thodoh1/FinancialPredictionJev · Python

    00
    jevpython
  • An evaluation of typesafe AI chess. As it turns out, the AI isn't doing really well even though chess is not a particularly open-ended game. Still, it's only a prototype and this probably wasn't optimzied for games.

    AliceRoselia/Typesafe_chess_eval · Python

    00
    jevpython
  • Experimental design harness: a small decision model (Jev) picks the design in about a second, a traditional LLM (Luna) only writes the words. With and without it.

    LamplighterPaul/forma-system1-experiment · TypeScript

    00
    jev
  • Decision library to detect and classify AI hallucinations, powered by Jev AI.

    zavocc/ground-zero · Python

    00
    jev
  • Does Jev predict stock returns from news? It reads the news well; there is no tradeable alpha. Three arms separate reading from recall.

    Gaurav-Gosain/jev-alpha-bench · Go

    00
    jevgo
  • Jev (TypeSafe) vs. Gemini 3.8 Flash vs. GPT-5.6 Luna na anotação estruturada de sentenças do TJSP: qualidade, tempo e custo

    lab-dados/jev-anotacao-sentencas · Python

    00
    jevpython
  • Challenge Atari with Jev: structured decisions, value questions, and replayable experiments

    memorysaver/jev-atari-lab · Python

    00
    jevdemo
  • TypeSafe / Jev community project: thomasschafer/jev-bench.

    thomasschafer/jev-bench · Python

    00
    jev
  • Position paper: the Hidden-Markov and fuzzy primitives missing from TypeSafe AI's Jev and System-One decision models. Two lemmas, one principle (Deferred Crispification), one architecture (BSF-S1).

    dnakhoa/jev-deferred-crispification · TeX

    00
    jevtexcalibrationfuzzy-logichidden-markov-model
  • Demos to test the effectiveness of TypeSafe's "Jev" System One Model

    Bud-ro/jev-demos · Dart

    00
    jevdart
  • 同じ発言を jev と LLM の両方に判定させ、感情の変動値のズレと応答速度を1画面で見比べるデモ(affectus + Vercel AI Gateway)

    n-yokomachi/jev-dev · TypeScript

    00
    jevtypescript
  • typesafe.ai model jev finance benchmark

    hifizz/jev-finance-benchmark

    00
    jev
  • Can Jev pick the winner of a real headline A/B test? 64.5% across 10,984 Upworthy randomized experiments, 74.7% when the difference was decisive.

    Gaurav-Gosain/jev-headline-bench · Go

    00
    jevgo
  • Jev (TypeSafe) 性能評価プロジェクト — 日本郵便 KEN_ALL をマスタに、AI SDK 経由の Jev が住所のあいまい一致にどこまで使えるかを検証

    smasato/jev-jp-address · TypeScript

    00
    jevtypescript
  • TypeScript experiments, evaluations, and latency benchmarks for TypeSafe's Jev model

    Menny1337/jev-lab · TypeScript

    00
    jevtypescript
  • A small reproducible MuJoCo pilot comparing Jev, Claude Haiku, and reactive rules for pick-and-place.

    tryaksh/jev-pick-and-place-study · Python

    00
    jevpython
  • 发明 RLHF 的人,这次做了个不会说话的模型:Jev 独立研究报告。52 页 PDF + 50 条中文实测复现包 + 143 条可回溯数据表

    HackSing/jev-report · Python

    00
    jevpythonai-researchchinesellm-evaluation
  • A small second eval for shadcn-ui/lint that uses TypeSafe's Jev to judge the linter's own output.

    blas0/jev-shadcn-lint-eval · JavaScript

    00
    jevjavascript
  • Application of TypeSafe Jev (noul judgment primitive) on the collusion.wiki corpus: agent vs human page authorship, head-to-head vs local Qwen3.8-Flash-Next

    sypherin/jev-trace-classifier · Python

    00
    jevpython
  • Reproducible Jev vs Luna review-classification benchmark with measured accuracy, latency, and costs.

    mameli/jev-vs-luna · Python

    00
    jev
  • A 3D planetary rover sandbox for experimenting with autonomous decisions using TypeSafe AI.

    juancamiloqhz/roverlab · TypeScript

    00
    jevtypescript
  • Evaluating TypeSafe's Jev as a fast monitor and action gate for agent sabotage in SHADE-Arena, compared with Gemini 2.5 Flash/Pro.

    nican2018/shade-arena-jev-monitor · Python

    00
    jevpython
  • A sandbox where a TypeSafe System One model presses the controls of a small creature. Code runs the world.

    TheGali/terrarium · JavaScript

    00
    jevjavascript
  • Charts: TypeSafe Jev evaluated on Thai standardized exams vs 110 other models

    vehas/thaiexam-jev-charts · HTML

    00
    jevhtml
  • TypeSafe / Jev community project: karimatayuta/tiny-jev.

    karimatayuta/tiny-jev · HTML

    00
    jev
  • Crypto trading bot on Binance testnet using TypeSafe (Jev) to judge news

    Spykoninho/trading-bot-jev · TypeScript

    00
    jev
  • Evaluating TypeSafe's System One primitives (Choice/Score/Noul) — where a typed oracle beats an LLM call

    trophee-bot/typesafe-oracles · JavaScript

    00
    jevjavascript
  • Small demos of Jev (TypeSafe) through the Vercel AI Gateway: wiki race, town of agents, bullet chess, and more

    az9713/jev-projects · JavaScript

    00
    jevjavascriptdemo
  • Historical paper-trading simulator for evaluating TypeSafe AI JEV decisions

    co1smos/jev-demo · Python

    00
    jevpythondemo
  • Jev (TypeSafe System One) と LLM に同じゲームを打たせて、レイテンシ・コスト・判断の質を比べる練習台

    ryuchan00/jev_practice · Python

    00
    jevpythonsystem-one
  • Does TypeSafe's Jev keep its accuracy and calibration on Russian? Independent RU vs EN audit (ECE, reliability diagrams, paired bootstrap) on parallel human-labelled data.

    AHTOOOXA/jev-cyrillic-audit · Python

    00
    jevpython
  • jev-llm-benchmark: TypeSafe Jev ecosystem repository.

    Chronona/jev-llm-benchmark · TypeScript

    00
    jevtypescript
  • Does a System One model actually beat keyword matching? A reproducible benchmark on catching disguised duplicate thesis titles. 24 cases, real production baseline, raw data and charts included.

    devnolife/jev-vs-tfidf-benchmark · TypeScript

    00
    jevbenchmarkindonesiainformation-retrievalllmplagiarism-detectionstructured-outputtf-idf
  • jev-btzsc: TypeSafe Jev ecosystem repository.

    Gazer2020/jev-btzsc · Python

    00
    jevpython
  • Evaluate typed AI decisions on labeled French-language cases.

    gbesse/jev-banc-francais · JavaScript

    00
    jevcalibrationevalsfrancetypesafe-aijavascript
  • Qualitative coding at scale with Jev: apply a codebook to open-ended text, review uncertain items, measure agreement against your human coders.

    gbesse/jev-codebook · Python

    00
    jevhuman-in-the-loopinter-rater-reliabilityopen-sourcepythonqualitative-researchtext-classification
  • Run a declared factorial audience grid through typed Jev reactions and expose disagreement.

    gbesse/jev-crowdsim · JavaScript

    00
    jevaudience-researchfactorial-designnodejsopen-sourcesimulationsynthetic-datajavascript
  • Score declared writing dimensions and cite only exact source spans for weak results.

    gbesse/jev-roast · JavaScript

    00
    jevcontent-qualityexplainable-ainodejsopen-sourcetext-evaluationwriting-analysisjavascript
  • Title and abstract screening for systematic reviews with Jev: explicit criteria, include/exclude/maybe with reasons, PRISMA counts, RIS export, recall against human screeners.

    gbesse/jev-screen · Python

    00
    jevevidence-synthesishuman-in-the-loopliterature-screeningopen-sourceprismapythonsystematic-review
  • Before every NFL snap, a decision-only AI model (TypeSafe Jev) calls run or pass and go/punt/kick on fourth down, graded live against the coach.

    gregjonesio/jev-nfl · JavaScript

    00
    jevjavascript
  • Jev (TypeSafe System One) 테스트베드 — 클라우드 API와 로컬 셀프호스팅(jeff/GLiFormer) 양쪽 실행 예제 및 실측 결과

    hulryung/jev-testbed · Python

    00
    jevpython
  • Local JEV experiments: route decisions, Minesweeper solvers, and drone simulation

    imom39a/jev-playground · Python

    00
    jevpython
  • Fifty real-world financial use cases for TypeSafe's Jev model: typed, structured LLM answers over ledgers, fraud, portfolios, trades and filings, each graded against data where the right answer is known.

    IslamBaraka90/jev-typesafe-real-financial-use-cases · JavaScript

    00
    jevai-agentsbacktestingfinancial-datafintechjavascriptllmllm-evaluation
  • A live, graphical dojo for TypeSafe's Jev (System One) typed decision model — routing, a Tetris-playing agent, parallel swarms, and an honest Jev-vs-Claude gauntlet.

    lafollett-labs/typesafe-jev-dojo · TypeScript

    01
    jevai-agentscanvasdecision-modelopenroutersystem-onetetristypesafe
  • Does your model's confidence mean anything on your data? Calibration layer for typed probabilistic decisions — reliability, ECE, Brier, recalibration maps and cost-aware thresholds from decisions + outcomes.

    Maher-Reven/calibrant · TypeScript

    00
    jevtypescript
  • ¿Jev entiende tu acento? Pre-registered audit of TypeSafe AI's Jev on Spanish — accuracy, calibration and token cost — plus a CLI to run the same comparison on your own labelled data.

    marcosmartinez/jev-acento · Python

    01
    jevbenchmarkcalibrationexpected-calibration-errorllm-evaluationnlppre-registrationreproducible-research
  • Measuring TypeSafe AI's Jev on 41-clause contract review (CUAD, 20,500 decisions) against fast, cheap LLMs — latency, cost and F1

    matu79go/jev-hanko · Python

    00
    jevbenchmarkcontract-reviewcuadlatencylegal-aillmopenrouter
  • Small experiments with Jev by TypeSafe

    nak1b/jev-experiments · TypeScript

    00
    jevaiai-experimentsjev-aitypesafe-aitypescript
  • demo trend seracher using jev

    nhchoi98/demo_trend_searcher · TypeScript

    00
    jevtypescript
  • Independent demo of TypeSafe's Jev model: typed decisions with probabilities, measured side by side with OpenAI on support-ticket triage. Live local app plus a recorded replay page.

    ogamircs/jev-demo · HTML

    00
    jevdemollm-benchmarkopenaistructured-outputstypesafe-aihtml
  • An 8-bit computer built from one yes/no question asked to Jev (TypeSafe) — 24,511 NAND gates from a single API call

    RiwRiwara/jev-computer · Python

    00
    jevpython
  • Independent playground for TypeSafe AI Jev decision models: typed decisions, support-ticket routing, reproducible evaluations, and a local browser demo.

    STiFLeR7/Jev-LLM-Playground · JavaScript

    00
    jevai-evaluationbenchmarkingdecision-modelsjavascriptnodejsplaygroundstructured-decisions
  • Tiny inputs. Instant decisions. A small experimental playground for exploring fast, probabilistic decisions with Jev.

    takafumikobayashi/jev-snap-lab · TypeScript

    00
    jevtypescript
  • Local Jev evaluation workbench: datasets, typed questions, threshold simulation and run comparison

    Tomdachs/jev-replay-lab · TypeScript

    00
    jevtypescript
  • english-2-sql: TypeSafe Jev ecosystem repository.

    trivektor/english-2-sql · JavaScript

    00
    jevjavascript
  • jeval: open-source evaluations for AI outputs and agents, judged by Jev

    vrash/jeval · TypeScript

    00
    jevtypescript
  • An independent, reproducible benchmark of TypeSafe's Jev against open, CPU-only alternatives — 10,000 decisions, all raw results published.

    zhlei07/open-system-one · Python

    01
    jevbenchmarkcpu-inferencecross-encoderembeddingstext-classificationtypesafezero-shot-classification
  • ask-twice: TypeSafe Jev ecosystem repository.

    aarongunasingh/ask-twice · Python

    00
    jevpython
  • A learning scaffold for TypeSafe AI's System One models: eval harness plus a measured, plain-language comparison of the Jev decision model vs an LLM stand-in on 24 real operational decisions. All numbers reproducible from committed run files.

    andreaserradev-gbj/jev-access-day · TypeScript

    00
    jevaibenchmarkdecision-supportevaluationllmtype-safetytypesafe-ai
  • jev-demo: TypeSafe Jev ecosystem repository.

    aoprisan/jev-demo · Rust

    00
    jevrust
  • Can a System One model play arcade games? TypeSafe's Jev plays Tetris, Snake and 2048 — benchmarked against random and heuristic baselines.

    CankatSarac/jev-arcade · Python

    00
    jevpython
  • Playground for TypeSafe's Jev System One model (Next.js)

    Dillettant/jev-test · TypeScript

    00
    jevtypescript
  • fraud-jev: TypeSafe Jev ecosystem repository.

    frankied003/fraud-jev · TypeScript

    00
    jevtypescript
  • A 60-game benchmark of TypeSafe's Jev evaluation model playing Battleship. The model matches plain code; it does not beat it.

    ickas/battleship-vs-jev · TypeScript

    00
    jevaibattleshipbenchmarkevaluation-metricsllm-evaluationtypescript
  • A cached Jev prior for active learning: reusable rank fusion, matched ASReview controls, and a no-key evidence replay.

    joaovaleri/jev-shortlist · Python

    00
    jevactive-learningasreviewpythonreciprocal-rank-fusionreproducible-researchsystematic-reviewtypesafe
  • JevScope: TypeSafe Jev ecosystem repository.

    KaushikKC/JevScope · TypeScript

    00
    jevtypescript
  • Jev × obniz LED: Physical AI Hello World — text → typed decisions (TypeSafe System One) → WS2812B LEDs

    kofujimura/jev-obniz-led · TypeScript

    00
    jevtypescript
  • jevmaze: TypeSafe Jev ecosystem repository.

    kt3k/jevmaze · TypeScript

    00
    jevtypescript
  • Real browser-agent safety evaluation: Jev versus a baseline on benign and injected tasks

    mjyoke1111/jev-lab · TypeScript

    00
    jevtypescript
  • Experimental PoC for Jev translation checking: bilingual benchmarks, prompt comparisons, re-verification, and MAGI voting.

    mshk/jev-translation-checker · JavaScript

    00
    jevjavascript
  • A pre-registered field trial of Jev (TypeSafe's judgment model) on a second brain and Claude Code history: 20 tests, bars written first, failures included, and the tools to repeat it.

    NaluKicks-808/jev-field-trial · Python

    00
    jevai-agentsclaude-codeevaluationsecond-braintypesafepython
  • Reproducible benchmark evaluating TypeSafe AI's Jev (System One paradigm) on Brazil's ENEM 2025 standardized exam. Evaluates typed decision-making, domain-specific accuracy, and RLCD uncertainty calibration against open LLM baselines with an interactive GitHub Pages dashboard.

    patryckalves/jev-no-enem · Python

    00
    jevbenchmarkbrasilenemllmrlcdsystem-one-modelspython
  • jev-integration-report: TypeSafe Jev ecosystem repository.

    piratchai/jev-integration-report

    00
    jev
  • Reproducible Tetris decision benchmark comparing TypeSafe Jev with Claude Haiku

    planstack-ai/jev-tetris-benchmark · TypeScript

    00
    jevai-evaluationnextjstetristypesafe-aivercel-ai-gatewaytypescript
  • WHAT-s-Up-jev: TypeSafe Jev ecosystem repository.

    Pragyan330/WHAT-s-Up-jev · Python

    00
    jevpython
  • 998jevy
    Jev-style typed-decision model distilled from official Jev. 118M, EN+CN, trains on a 4GB GPU in 10 minutes.

    slatinwine/jevy · Python

    00
    jevpython
  • jev-noul-vs-choice: TypeSafe Jev ecosystem repository.

    TakumiNoguchi2004/jev-noul-vs-choice · Python

    00
    jevpython
  • TypeSafe AI の判定モデル jev に 2048 を遊ばせる PoC(Go CLI + Cloudflare Workers の Web デモ)

    tatsuo48/jev-poc · Go

    00
    jevgo
  • 997jevx
    Jev (TypeSafe System One) research: API notes, benchmarks, community experiments, agent-loop patterns

    umgbhalla/jevx · Python

    00
    jevpython
  • 732jev
    Independent research notes toward an open Jev-like decision model: public facts, API contract, training and eval plan.

    WiredMind2/jev · Python

    00
    jevpython
  • Experiments with TypeSafe's Jev System One model

    xavierforge/jev_experiments · JavaScript

    00
    jevjavascript
  • Live demo showing why loop-speed classification matters: Jev vs LLMs on the same events, same clock, honest scorecard.

    yshraj/jev-traffic-race · TypeScript

    00
    jevtypescript
  • Label every sentence of a document with calibrated probabilities from Jev, rendered as a heatmap

    abhishekmishragithub/semantic-microscope · Python

    00
    jevllmpythontypesafevisualization
  • Auto-marking maths scripts with Jev (TypeSafe System One): 2,054 scripts, 96.6% agreement with human markers

    Akeel-Majeed/JEValuate · TypeScript

    00
    jevtypescript
  • security-sandbox-jev: TypeSafe Jev ecosystem repository.

    altanapps/security-sandbox-jev · Python

    00
    jevpython
  • 1020Mirave
    Mirave: TypeSafe Jev ecosystem repository.

    edoigtrd/Mirave · Python

    00
    jevpython
  • Training demonstration of Jev in a dispatch services command center scenario.

    fullcolorcoder/reflex-jev · TypeScript

    00
    jevtypescript
  • Empirical experiments and API research for TypeSafe's Jev System One model trained using RLCD

    BipinRajC/Jev-api-experiments · Python

    00
    jevapiexperimentalrlcdtypesafe-aipython
  • jev-experiments: TypeSafe Jev ecosystem repository.

    harlanljones/jev-experiments · JavaScript

    00
    jevjavascript
  • jev-experiments: TypeSafe Jev ecosystem repository.

    pavan142/jev-experiments · TypeScript

    00
    jevtypescript