Fold hourly 2246 HIGH - #58
Conversation
Densify Open-Jev TREC prep on §125 and TypeLLM PyPI 0.1.1 on §113. First-card simple-jev plus directories, memory, LoRA retrospective, serving packs, and namesakes. TREC prep is not completed TREC. logits are not calibrated probabilities of correctness. Co-authored-by: Basit Mustafa <24601@users.noreply.github.com>
|
Testing gate: PASS Independent adversarial review + tests on head Prior twenty uniqueness locks (0843–2146 + Open-Jev densify) + REVISIT_LOCK are byte-identical vs
IDs. Claimed Adversarial
Notes, not FAILs. Star-noise after snapshot: TypeLLM 24★ vs fold 23★; simple-jev 409★ vs fold 408★. Open-Jev 19★ still matches. Hub archerhume/4rcherhume HTTP 404 this pass (fold 401). Description/README gitblobs still match. Prompt SHA TypeLLM Ready for parent review. PR stays open. Do not merge from this agent. |
Off latest
mainafter merged #57 (f635ffeFold hourly 2146 HIGH). Do not reopen or amend PR #23–#57. Do not bump 0.5.0.Hourly 2246 HIGH fold. Jev is the exemplar of a backend-agnostic class of categorization/scoring/typed-decision models, not the whole mandate. REVISIT densifies original cards:
Zefan-Cai/Open-Jevon §125,TypeLLM/TypeLLMon §113.featherless-ai/simple-jevis a first card this hour (jevinize stays a pointer). Do not mint sibling first sightings. SHA move is not a replica. Third-party benches stay theirs.REVISIT densify (theirs)
Zefan-Cai/Open-JevTREC evaluation preparation and context proof. Live 19★ (star-noise is not the fold). HEADa00559ea0ab2→48346d0630f1(Publish strict Open-Jev TREC evaluation preparation and context proof). README SHA unchangedce1a587219e4. Quote theirs: Actual Open-Jev TREC model inference is pending; All 79 combined CPU tests pass; 97 queries 43 DL19 54 DL20; at most 873 requests per model; No GPU or model inference was used. TREC prep ≠ completed Open-Jev TREC. context proof ≠ nDCG. CPU tests ≠ GPU scores. Open-Jev TREC pending.TypeLLM/TypeLLMPyPI packaging. Live 23★. Prompt SHA702e6a287f3c→624460ed65fb(banner/CI); live HEAD8a8b4aefd443(Release 0.1.1). README SHA9f6dea3a4c8c. Quote theirs: Add PyPI packaging and publish workflow; typellm 0.1.1; Drop fixed banner height so it scales on PyPI. Constrained AR ≠ calibrated Noul. PyPI packaging ≠ calibrated Noul. type safety does not guarantee factual accuracy.Novel HIGH (first-sighting)
featherless-ai/simple-jev408★ first card. Turn any open model into a classifier/jev endpoint. logits are not calibrated probabilities of correctness. does not reproduce TypeSafe./v1/systemonealias of/v1/classifier. wire-compat ≠ logit-equiv.everyai-com/jev-directory13★. 50 runnable evals / 1300+ builds. catalog ≠ endorsement.libingzheren/Jev-Mem11.0% 6.6× 36.7% theirs not Harbor.Nyarlathoteppppp/pi-jev-context≠ kevinpita/pi-jev-context. Cache-neutral trim.FogMoe/necroabandoned LoRA retrospective. LoRA ≠ RLCD replica. Qwen3.5-0.8B ≠ Archer.TypeSafeAI/clarity-judgeindependent community project. Simulated demo is not measured Jev.ziozzang/hearimJev-compatible Go gateway.yijunyu/jev-rsany LLM one prefill.Skip Archer (
promised_not_landed). Hub archerhume/4rcherhume HTTP 401.IDs
notes.md§134 (densify cards stay on §125 / §113)Class recipes (not Jev-only)
Without Augustus: treat a CPU-passing TREC prep as completed Open-Jev nDCG, a PyPI wheel as a calibrated Noul, next-token logits as P(correct), a 50-eval directory as endorsement, or LoCoMo 11.0% as Harbor.
With Augustus: TREC prep ≠ completed Open-Jev TREC; context proof ≠ nDCG; CPU tests ≠ GPU scores; PyPI packaging ≠ calibrated Noul; Constrained AR ≠ calibrated Noul; logits are not calibrated probabilities of correctness; wire-compat ≠ logit-equiv; catalog ≠ endorsement; theirs not Harbor. Same split for any Choice/Score/Noul-style head.
Honesty: theirs on bench numbers; catalog ≠ endorsement; soft judgment never sole veto.
Does not bump 0.5.0. Skip Archer.
invented_signal: false.