davwer commited on
Commit
edfa877
·
verified ·
1 Parent(s): e4506b8

Upload TM-NLA release checkpoints

Browse files
README.md CHANGED
@@ -1,3 +1,33 @@
1
- ---
2
- license: mit
3
- ---
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ # TM-NLA Checkpoints
2
+
3
+ These checkpoints support the TM-NLA public runtime.
4
+
5
+ TM-NLA is a proof-of-concept temporal activation-verbalization probe for frozen visual-temporal hidden states. It projects video-window activations from `Qwen/Qwen3.5-0.8B` into a continuous language-oriented manifold, then selects semantic readout hypotheses over time with compatibility and specificity annotations.
6
+
7
+ ## Files
8
+
9
+ | File | Role | SHA-256 |
10
+ | --- | --- | --- |
11
+ | `temporal_probe.pt` | Temporal visual activation -> continuous language-manifold probe | `48AABF662423EA529C985855EA54970A9CE3256EC3F14AFE0169C2BF1A55B217` |
12
+ | `activation_verbalizer.pt` | Activation -> small set of short candidate readout proposals | `88DDB0267AE535A287755DE5C9E638EB771FFD5C5377051ACFCADF96A98C08C1` |
13
+ | `text_reconstructor.pt` | Text readout -> activation compatibility verifier/reranker | `C5CEA721A2173DCA884617258BB3F671668CF13EBA07728E11C39C9695B5DE32` |
14
+
15
+ ## Base Model
16
+
17
+ The runtime uses frozen `Qwen/Qwen3.5-0.8B` activations. The public demo path does not train or fine-tune Qwen.
18
+
19
+ ## Intended Use
20
+
21
+ Use these checkpoints with the TM-NLA GitHub repository to run local terminal-based semantic readouts over video windows.
22
+
23
+ The checkpoints are intended for interpretability and research exploration. They are not intended as a production captioner, classifier, benchmark system, or broad video-analysis product.
24
+
25
+ ## Limitations
26
+
27
+ The continuous activation geometry is the strongest result. TM-NLA shows weak recoverable temporal semantic traces from frozen visual-temporal activations. Natural-language readouts are useful but noisy externalizations, not ground-truth captions.
28
+
29
+ Readouts can be malformed, wrong, generic, or underspecified. The activation verbalizer remains fragile, and candidate generation plus AR reranking is a practical readout mechanism rather than a solved decoder. The text reconstructor is a compatibility-based verifier/reranker, not an oracle. `specificity_status` is a numerical activation-specificity annotation, not a ground-truth correctness label.
30
+
31
+ ## Citation
32
+
33
+ If you publish work building on this proof of concept, cite or link the companion TM-NLA GitHub repository.
activation_verbalizer.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:88ddb0267ae535a287755de5c9e638eb771ffd5c5377051acfcadf96a98c08c1
3
+ size 12627790
temporal_probe.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:48aabf662423ea529c985855ea54970a9ce3256ec3f14afe0169c2bf1a55b217
3
+ size 163736830
text_reconstructor.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:c5cea721a2173dca884617258bb3f671668cf13eba07728e11c39c9695b5de32
3
+ size 12587104