Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🧙
Computational semiotics is empirical.
Burton Lancaster
PRO
RiverRider
5
3
17
Follow
keblezhu's profile picture
3nesdeniz's profile picture
redfastner's profile picture
32 followers
·
33 following
https://sunstonenorth.com
Space-Bacon
AI & ML interests
Explainable AI
Recent Activity
updated
a dataset
about 6 hours ago
RiverRider/srt-omni-crossvendor-states
posted
an
update
about 10 hours ago
A 339 KB linear probe on frozen features beats the fine-tuned baseline on ChestX-ray14. Linear(5376, 14) on frozen google/gemma-4-31B-it hidden states. No fine-tuning, no radiology pretraining, no augmentation. All 112,120 images, official test_list.txt. Wang et al. 2017, ResNet-50 fine-tuned end to end 0.7451 this probe, frozen backbone + linear head 0.7590 view-position only (shortcut baseline) 0.5896 shuffled labels (refit floor) 0.5002 Ahead on 12 of 14 findings. The comparison is split-matched, and that took care to get right. The number everyone quotes, CheXNet's 0.8414, is on a different test set: their own random 70/10/20 partition, not the official list. Do not compare 0.7590 to it. The matched row is from Wang's v5 appendix, added specifically to report the published split. I had this wrong in our own code for a day, quoting a cross-split reference as a head-to-head, which is the error worth not repeating in public. Three controls, because a bare AUROC here is not interpretable. Shuffled labels catch leakage. View-only catches the shortcut, since portable AP films are taken of sicker patients, and it is folded, because Hernia's raw view-only of 0.3436 is really 0.6564 of shortcut once flipped. Intervals resample patients and not images, since the test split is 25,596 films from 2,797 patients. Banked negatives are on the card too. Max-pooling and top-16 pooling were predicted to help focal findings and did the opposite, costing 0.0537 and 0.0225. Readout depth barely matters, 0.7600 to 0.7605. Scope: detection, not early detection. Research artifact, not a diagnostic device. The backbone never runs in the demo. What ships is the reading. Space: https://huggingface.co/spaces/RiverRider/srt-cxr14-probe Model: https://huggingface.co/RiverRider/srt-cxr14-linear-probe Data + states: https://huggingface.co/datasets/RiverRider/srt-cxr14-frozen-probe
updated
a collection
about 10 hours ago
SRT demos
View all activity
Organizations
RiverRider
's datasets
12
Sort: Recently updated
RiverRider/srt-omni-crossvendor-states
Updated
about 6 hours ago
•
59
RiverRider/srt-cxr14-frozen-probe
Updated
about 11 hours ago
RiverRider/srt-nla-gemma4-artifacts
Updated
1 day ago
•
71
RiverRider/srt-depth-probe-artifacts
Updated
1 day ago
•
43
RiverRider/srt-omni-manifest
Viewer
•
Updated
2 days ago
•
1
•
20
RiverRider/srt-qwen38-coco-states
Updated
6 days ago
•
35
RiverRider/srt-coco-thumbs
Viewer
•
Updated
6 days ago
•
10.3k
•
369
RiverRider/srt-nla-gptoss20b-artifacts
Updated
Jul 2
•
108
RiverRider/srt-nla-targets-gemma2-2b-v1
Updated
May 21
•
26
RiverRider/srt-nla-targets-llama32-3b-v1
Updated
May 21
•
21
RiverRider/srt-nla-targets-v1
Updated
May 18
•
44
RiverRider/zoolander-corpus-v23
Viewer
•
Updated
May 5
•
93.3k
•
52