DNA Foundation Model Benchmark
Compare DNA language models across genomics classification tasks.
Drill into model variants
Pick one model, then optionally isolate a single config dimension below.
Holds the other two dimensions fixed at their best combo, so bars only differ by this one.
Drill into individual tasks
Each point = one variant (mean metric across all tasks vs mean speed across tasks). Pareto frontier = variants not dominated on both axes (black outline).
Efficiency ranking
Variants sorted by performance โ speed and derived efficiency metrics
1 | 2 | 3 |
|---|---|---|
Applies immediately to all charts โ no need to click Apply Filters below.
Currently returns no data โ all runs use subsample_train=0.05.
Exclude specific run IDs
Preview (first 50 rows)
1 | 2 | 3 |
|---|---|---|
1 | 2 | 3 |
|---|
Select a model and embedding config to audit which runs were kept or dropped by dedup.
Runs for selection (โ = kept by dedup, โ = dropped)
1 | 2 | 3 |
|---|---|---|
1 | 2 | 3 |
|---|