Team Ai
Modelpublic

BaseIntelligence/top-prism-architecture

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
0likes72downloads
Model Card

<div align="center">

BASE Banner

<h1 align="center">PRISM top architecture</h1>

<p align="center"><b>Global-best miner architecture on Base PRISM — benchmarks vs GPT-2 / GPT-2 Large</b></p>

</div>


Benchmarks vs GPT-2 (Prism-protocol)

Prism-protocol public eval pack (1×RTX 5090). Accuracy: ↑ higher better. BPB: ↓ lower better. References: GPT-2 (124M) · GPT-2 Large (774M) (eval-only; not miner trains).

MetricThis modelGPT-2GPT-2 Largevs GPT-2vs GPT-2 Large
Val BPB (G1)3.71814.75954.1639↓ -1.0414 ✓ better↓ -0.4458 ✓ better
HellaSwag0.3600.3550.395↑ +0.005 ✓ better↓ -0.035 worse
ARC-Easy0.3350.2450.280↑ +0.090 ✓ better↑ +0.055 ✓ better
ARC-Challenge0.2950.2400.280↑ +0.055 ✓ better↑ +0.015 ✓ better
PIQA0.6300.5850.690↑ +0.045 ✓ better↓ -0.060 worse
WinoGrande0.5200.5150.545↑ +0.005 ✓ better↓ -0.025 worse
BoolQ0.6300.5750.640↑ +0.055 ✓ better↓ -0.010 worse
LAMBADA0.9550.9700.985↓ -0.015 worse↓ -0.030 worse
OpenBookQA0.3100.3200.335↓ -0.010 worse↓ -0.025 worse

Compute notes

This modelGPT-2GPT-2 Large
Parameters107.0M124M774M
Size vs Large7.23× vs GPT-2 Large (774M)6.22×1×
Train tokens—(eval-only)(eval-only)
Wall clock20539s(eval-only)(eval-only)
Sustained train throughput—n/an/a
GPU (harness)GPU 0: NVIDIA GeForce RTX 5090 (UUID: GPU-e31dbb89-6a01-a2fb-2684-3b6f0efb3f28)1×RTX 5090 (eval)1×RTX 5090 (eval)

Throughput ≈ 6 × N × D / wall TFLOPS (dense transformer train FLOPs rule of thumb).

Model card

fieldvalue
arch_idarch_f17d92b32a8c79f7
bpb3.718067
submission7b8658aedfdb158782567d09e26a9e819c578ff20c4db1d940e12de77cd9d6d2
owner_hotkey462c4a7dfe30…
hub repoBaseIntelligence/top-prism-architecture

Load (trustremotecode)

python
from transformers import AutoModel, AutoConfig
cfg = AutoConfig.from_pretrained("BaseIntelligence/top-prism-architecture", trust_remote_code=True)
model = AutoModel.from_pretrained("BaseIntelligence/top-prism-architecture", trust_remote_code=True)

Weights: checkpoint.pt (Hub LFS when large). Load via PrismCustomModel.from_pretrained with trust_remote_code=True.

Companion GitHub publish (when configured) lives under BaseIntelligence/prism top-model/.