Paper deep dive
Self-Guided Test-Time Training for Long-Context LLMs
Xinyu Zhu, Zhe Xu, Xiaohan Wei, Yunchen Pu, Fei Tian, Chonglin Sun, Kaushik Rangadurai, Hua Zhi, Frank Shyu, Sandeep Pandey, Luke Simon, Yu Meng, Xi Liu
Intelligence
Status: succeeded | Model: Gemma-4-26B-A4B | Prompt: intel-v1 | Confidence: 94%
Last extracted: 7/18/2026, 3:32:18 PM
Summary
The paper introduces Self-Guided Test-Time Training (S-TTT), a method for improving long-context reasoning in Large Language Models (LLMs). It addresses the issue that standard Test-Time Training (TTT) on full or randomly sampled contexts introduces noise and degrades performance. S-TTT uses the LLM to identify relevant evidence spans within the long context before adaptation, applying training only to these high-quality spans. This approach significantly improves accuracy on benchmarks like LongBench-v2 and LongBench-Pro for models such as Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct.
Entities (8)
Relation Signals (7)
Self-Guided Test-Time Training → appliedto → Llama-3.1-8B-Instruct
confidence 96% · S-TTT improves accuracy for both Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct
Self-Guided Test-Time Training → appliedto → Qwen3-4B-Thinking-2507
confidence 96% · S-TTT improves accuracy for both Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct
Self-Guided Test-Time Training → improves → LongBench V2
confidence 95% · S-TTT improves accuracy for both Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct... On two challenging long-context reasoning benchmarks, LongBench-v2 and LongBench-Pro
Self-Guided Test-Time Training → improves → LongBench-Pro
confidence 95% · S-TTT improves accuracy for both Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct... On two challenging long-context reasoning benchmarks, LongBench-v2 and LongBench-Pro
Meta AI → affiliationof → Xinyu Zhu
confidence 90% · Xinyu Zhu 1,2 ... 1 Meta AI
Self-Guided Test-Time Training → mitigates → noise
confidence 90% · mitigating the severe noise of random span sampling
Test-Time Training → suffersfrom → noise
confidence 90% · adapting on randomly sampled spans introduces severe noise
Cypher Suggestions (0)
No Cypher suggestions yet.
Abstract
Abstract:Long-context processing has become increasingly important for large language models (LLMs), but simply extending the context window does not guarantee effective utilization of long inputs. As input length grows, accuracy often degrades, indicating that models still struggle to identify and use the evidence most relevant to a question. A promising way to improve long-context utilization is test-time training (TTT), which treats the test context as a training example for instance-specific parameter adaptation. However, applying TTT to the entire long context is prohibitively expensive, while adapting on randomly sampled spans introduces severe noise. Because most spans in a long context are irrelevant to the specific question, training on them may even degrade the base model's performance. Our preliminary study shows that TTT is highly sensitive to training-span quality: on LongBench-v2, TTT on randomly sampled spans hurts performance, whereas TTT on oracle spans substantially improves it. Motivated by this, we propose a simple method, Self-Guided TTT (S-TTT): before adaptation, the model identifies the evidence spans it should learn from, and the standard language-modeling training objective is applied only to those selected spans. On two challenging long-context reasoning benchmarks, LongBench-v2 and LongBench-Pro, S-TTT improves accuracy for both Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct, achieving up to a 15% relative improvement.
Tags
Links
- Source: https://arxiv.org/abs/2607.09415v1
- Canonical: https://arxiv.org/abs/2607.09415v1
Trouble viewing inline? Open PDF directly →
Full Text
100,330 characters extracted from source content.
Expand or collapse full text
Self-Guided Test-Time Training for Long-Context LLMs Xinyu Zhu 1,2 , Zhe Xu 1,† , Xiaohan Wei 1 , Yunchen Pu 1 , Fei Tian 1 , Chonglin Sun 1 , Kaushik Rangadurai 1 , Hua Zhi 1 , Frank Shyu 1 , Sandeep Pandey 1 , Luke Simon 1 , Yu Meng 2,‡ , Xi Liu 1,‡ 1 Meta AI, 2 University of Virginia † Execution Lead, ‡ Joint corresponding author Long-context processing has become increasingly important for large language models (LLMs), but simply extending the context window does not guarantee effective utilization of long inputs. As input length grows, accuracy often degrades, indicating that models still struggle to identify and use the evidence most relevant to a question. A promising way to improve long-context utilization is test-time training (T), which treats the test context as a training example for instance-specific parameter adaptation. However, applying T to the entire long context is prohibitively expensive, while adapting on randomly sampled spans introduces severe noise. Because most spans in a long context are irrelevant to the specific question, training on them may even degrade the base model’s performance. Our preliminary study shows that T is highly sensitive to training-span quality: on LongBench-v2, T on randomly sampled spans hurts performance, whereas T on oracle spans substantially improves it. Motivated by this, we propose a simple method, Self-Guided T (S-T): before adaptation, the model identifies the evidence spans it should learn from, and the standard language-modeling training objective is applied only to those selected spans. On two challenging long- context reasoning benchmarks, LongBench-v2 and LongBench-Pro, S-T improves accuracy for both Qwen3-4B-Thinking-2507 and Llama-3.1-8B-Instruct, achieving up to a 15% relative improvement. Correspondence: Yu Meng (yumeng5@virginia.edu) and Xi Liu (xliu1@meta.com) Date: July 10, 2026 1 Introduction Long-context capability has become a central requirement for modern language models. Recent models support context windows of hundreds of thousands of tokens, enabling them to process long inputs in a single prompt (Peng et al., 2024; Chen et al., 2024). Despite this progress, a larger window does not by itself ensure that the model can use long inputs effectively. As context length grows, accuracy often degrades, and models struggle to keep the most relevant evidence accessible throughout reasoning and decoding (Liu et al., 2024; Hsieh et al., 2024). This suggests that the bottleneck in long-context reasoning is not merely fitting more tokens into the prompt, but ensuring that the model can identify and use the evidence relevant to the question. Test-time training (T) (Sun et al., 2020; Liu et al., 2021; Hardt and Sun, 2024; Akyürek et al., 2024; Tandon et al., 2025; Zhang et al., 2025a; Feng et al., 2026) has emerged as a promising solution. Instead of answering with a fixed model, T treats the test input itself as a training example, adapts the model weights for that specific instance, and uses the adapted weights to generate the answer. For long-context tasks, this is especially appealing because adaptation can turn instance-specific evidence in the context into parameter updates, making it easier to use during subsequent generation (Bansal et al., 2026; Chen et al., 2026a). However, a key challenge in applying T to long contexts is determining what data to train on—an important dimension that remains largely underexplored. Existing approaches commonly rely on either full-context adaptation (Tandon et al., 2025; Zhang et al., 2025a) or randomly sampled training spans(Bansal et al., 2026), both of which suffer from noisy signals. Not only is performing T on the full context computationally expensive, but it also overwhelms the adaptation process with distractors, as the vast majority of a long context is usually irrelevant to the specific query. A cheaper alternative is to train on randomly sampled 1 arXiv:2607.09415v1 [cs.CL] 10 Jul 2026 spans. While this mitigates the computational cost, it may amplify the noise: random sampling frequently misses the relevant evidence, causing the model to adapt primarily on distractors. This suggests that the central bottleneck of long-context T is not the adaptation mechanism itself, but rather test-time training-data quality. We empirically demonstrate this sensitivity through a preliminary diagnostic: on LongBench-v2 (Bai et al., 2025), T on random spans slightly degrades performance relative to standard base model inference, whereas training on answer-aware oracle spans annotated by GPT-5.5 yields substantial improvements. This demonstrates that the effectiveness of T depends critically on the signal-to-noise ratio of the training tokens. Motivated by this insight, we propose a simple solution, Self-Guided T (S-T). Rather than processing the entire context or sampling spans blindly, S-T leverages the LLM itself as a test-time data selector. We prompt the model to mark verbatim spans in the context that are likely to support the question. We then adapt the model on the selected spans with a next-token-prediction objective and generate the final answer from the full context. As such, S-T leaves the training objective, model architecture, and final decoding procedure unchanged; it optimizes only the test-time tokens used for adaptation. On two challenging long-context reasoning benchmarks, LongBench-v2 (Bai et al., 2025) and LongBench-Pro (Chen et al., 2026b), using Qwen3-4B-Thinking-2507 (Qwen Team, 2025) and Llama-3.1-8B-Instruct (Team, 2024) models, S-T consistently improves long-context performance and outperforms strong T baselines. Our contributions are: 1.We identify training-data quality as a critical yet underexplored bottleneck for long-context T. We empirically demonstrate that adapting on noisy context can degrade performance, whereas high-quality evidence spans lead to substantial gains. 2. We propose Self-Guided T (S-T), a simple and effective framework that uses the LLM itself to select question-relevant evidence spans for test-time training, avoiding the expensive computational cost of full-context training and mitigating the severe noise of random span sampling. 3.We evaluate S-T on two challenging long-context reasoning benchmarks LongBench-v2 and LongBench- Pro using Qwen3-4B-Thinking and Llama-3.1-8B-Instruct models. S-T consistently improves long- context performance and outperforms various strong T baselines. 2 Method 2.1 Preliminary analysis Test-time training has been used to improve LLMs long-context performance by adapting the model to the specific context observed at test time. For long inputs, however, directly training on the full context is expensive, and a naive alternative is to train on short spans sampled uniformly from the context. This reduces the compute but has a cost: in a long document, most uniformly sampled spans are irrelevant to the question. As a result, T on randomly sampled spans may adapt the model to distractors rather than evidence. MethodLongBench-v2 Base Model40.4 Random Span T38.9 Oracle Span T45.9 Table 1 Test-time training is sensitive to training-token quality. Training Qwen3-4B-Thinking-2507 on random span tokens does not lead to improvement; instead, it hurts performance. Table 1 shows a diagnostic experiment on LongBench-v2. The base model Qwen3-4B-Thinking-2507 reaches 40.4% accuracy without fine-tuning. After adapting the model via T on uniformly sampled spans, accuracy drops to 38.9%, indicating that T does not guarantee an improvement when the training tokens are noisy. 2 <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · Stage 1: Model-Guided Span Selection <latexit sha1_base64="0qSiA+TXOQEOvRSqdJWQbG5jZ2U=">AAAB6nicbVDLTgJBEOzFF+IL9ehlIjHxRHaJQY9ELx4xyiOBDZkdemHC7OxmZtZICJ/gxYPGePWLvPk3DrAHBSvppFLVne6uIBFcG9f9dnJr6xubW/ntws7u3v5B8fCoqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLRzcxvPaLSPJYPZpygH9GB5CFn1Fjp/qlX6RVLbtmdg6wSLyMlyFDvFb+6/ZilEUrDBNW647mJ8SdUGc4ETgvdVGNC2YgOsGOppBFqfzI/dUrOrNInYaxsSUPm6u+JCY20HkeB7YyoGeplbyb+53VSE175Ey6T1KBki0VhKoiJyexv0ucKmRFjSyhT3N5K2JAqyoxNp2BD8JZfXiXNStmrlqt3F6XadRZHHk7gFM7Bg0uowS3UoQEMBvAMr/DmCOfFeXc+Fq05J5s5hj9wPn8AEMaNrA==</latexit> x 2 <latexit sha1_base64="P4GArwQ3oaXEfJT4toixLcIzE8A=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qnn9Uplt+LOQJaJl5My5Kj3Sl/dfszSiCtkkhrT8dwE/YxqFEzySbGbGp5QNqID3rFU0YgbP5udOiGnVumTMNa2FJKZ+nsio5Ex4yiwnRHFoVn0puJ/XifF8MrPhEpS5IrNF4WpJBiT6d+kLzRnKMeWUKaFvZWwIdWUoU2naEPwFl9eJs3ziletVO8uyrXrPI4CHMMJnIEHl1CDW6hDAxgM4Ble4c2Rzovz7nzMW1ecfOYI/sD5/AEPQo2r</latexit> x 1 <latexit sha1_base64="7fJgorBTuC0IU+mE+H26FyrGvEQ=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qlneqWyW3FnIMvEy0kZctR7pa9uP2ZpxBUySY3peG6CfkY1Cib5pNhNDU8oG9EB71iqaMSNn81OnZBTq/RJGGtbCslM/T2R0ciYcRTYzoji0Cx6U/E/r5NieOVnQiUpcsXmi8JUEozJ9G/SF5ozlGNLKNPC3krYkGrK0KZTtCF4iy8vk+Z5xatWqncX5dp1HkcBjuEEzsCDS6jBLdShAQwG8Ayv8OZI58V5dz7mrStOPnMEf+B8/gBzSo3t</latexit> x s <latexit sha1_base64="v7xFkNYPbfbModP8QyvO1wHTzxo=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120i7dbMLuRiyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEsG1cd1vZ2V1bX1js7BV3N7Z3dsvHRw2dZwqhg0Wi1i1A6pRcIkNw43AdqKQRoHAVjC6mfqtR1Sax/LBjBP0IzqQPOSMGivdP/WwVyq7FXcGsky8nJQhR71X+ur2Y5ZGKA0TVOuO5ybGz6gynAmcFLupxoSyER1gx1JJI9R+Njt1Qk6t0idhrGxJQ2bq74mMRlqPo8B2RtQM9aI3Ff/zOqkJr/yMyyQ1KNl8UZgKYmIy/Zv0uUJmxNgSyhS3txI2pIoyY9Mp2hC8xZeXSfO84lUr1buLcu06j6MAx3ACZ+DBJdTgFurQAAYDeIZXeHOE8+K8Ox/z1hUnnzmCP3A+fwBeEo3f</latexit> x e <latexit sha1_base64="HAVxy/77DfEejtoQbW/WOeu2/bM=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120i7dbOLuRiihP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEsG1cd1vZ2V1bX1js7BV3N7Z3dsvHRw2dZwqhg0Wi1i1A6pRcIkNw43AdqKQRoHAVjC6mfqtJ1Sax/LBjBP0IzqQPOSMGivdP/a8XqnsVtwZyDLxclKGHPVe6avbj1kaoTRMUK07npsYP6PKcCZwUuymGhPKRnSAHUsljVD72ezUCTm1Sp+EsbIlDZmpvycyGmk9jgLbGVEz1IveVPzP66QmvPIzLpPUoGTzRWEqiInJ9G/S5wqZEWNLKFPc3krYkCrKjE2naEPwFl9eJs3ziletVO8uyrXrPI4CHMMJnIEHl1CDW6hDAxgM4Ble4c0Rzovz7nzMW1ecfOYI/sD5/AEEmI2k</latexit> q 1 <latexit sha1_base64="qb/tfXDey1ZHG0Gq1YoCB5K5dGA=">AAAB6nicbVDLTgJBEOzFF+IL9ehlIjHxRHaJQY9ELx4xyiOBDZkdZmHC7Ow602tCCJ/gxYPGePWLvPk3DrAHBSvppFLVne6uIJHCoOt+O7m19Y3Nrfx2YWd3b/+geHjUNHGqGW+wWMa6HVDDpVC8gQIlbyea0yiQvBWMbmZ+64lrI2L1gOOE+xEdKBEKRtFK94+9Sq9YcsvuHGSVeBkpQYZ6r/jV7ccsjbhCJqkxHc9N0J9QjYJJPi10U8MTykZ0wDuWKhpx40/mp07JmVX6JIy1LYVkrv6emNDImHEU2M6I4tAsezPxP6+TYnjlT4RKUuSKLRaFqSQYk9nfpC80ZyjHllCmhb2VsCHVlKFNp2BD8JZfXiXNStmrlqt3F6XadRZHHk7gFM7Bg0uowS3UoQEMBvAMr/DmSOfFeXc+Fq05J5s5hj9wPn8ABhyNpQ==</latexit> q 2 <latexit sha1_base64="iANey9rYlpgkCfH1CXEhuEbhg2o=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSIIQklEqseiF48V7Ae0oWy2k3bpZhN2N2IJ/RFePCji1d/jzX/jts1BWx8MPN6bYWZekAiujet+Oyura+sbm4Wt4vbO7t5+6eCwqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLR7dRvPaLSPJYPZpygH9GB5CFn1Fip9dTL9Lk36ZXKbsWdgSwTLydlyFHvlb66/ZilEUrDBNW647mJ8TOqDGcCJ8VuqjGhbEQH2LFU0gi1n83OnZBTq/RJGCtb0pCZ+nsio5HW4yiwnRE1Q73oTcX/vE5qwms/4zJJDUo2XxSmgpiYTH8nfa6QGTG2hDLF7a2EDamizNiEijYEb/HlZdK8qHjVSvX+sly7yeMowDGcwBl4cAU1uIM6NIDBCJ7hFd6cxHlx3p2PeeuKk88cwR84nz8Rm49p</latexit> x s+1 <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · Question Tokens <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="kfyhpuCjonc2/DMNNayB027cIxc=">AAAB6nicbVDLSgNBEOyNrxhfUY9eBoPgKeyKRI9BLx4j5gXJEmYns8mQ2dllplcMIZ/gxYMiXv0ib/6Nk2QPmljQUFR1090VJFIYdN1vJ7e2vrG5ld8u7Ozu7R8UD4+aJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LR7cxvPXJtRKzqOE64H9GBEqFgFK308NSr94olt+zOQVaJl5ESZKj1il/dfszSiCtkkhrT8dwE/QnVKJjk00I3NTyhbEQHvGOpohE3/mR+6pScWaVPwljbUkjm6u+JCY2MGUeB7YwoDs2yNxP/8zophtf+RKgkRa7YYlGYSoIxmf1N+kJzhnJsCWVa2FsJG1JNGdp0CjYEb/nlVdK8KHuVcuX+slS9yeLIwwmcwjl4cAVVuIMaNIDBAJ7hFd4c6bw4787HojXnZDPH8AfO5w9ETo3O</latexit> x T Base LLM Long Context Tokens <latexit sha1_base64="7fJgorBTuC0IU+mE+H26FyrGvEQ=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qlneqWyW3FnIMvEy0kZctR7pa9uP2ZpxBUySY3peG6CfkY1Cib5pNhNDU8oG9EB71iqaMSNn81OnZBTq/RJGGtbCslM/T2R0ciYcRTYzoji0Cx6U/E/r5NieOVnQiUpcsXmi8JUEozJ9G/SF5ozlGNLKNPC3krYkGrK0KZTtCF4iy8vk+Z5xatWqncX5dp1HkcBjuEEzsCDS6jBLdShAQwG8Ayv8OZI58V5dz7mrStOPnMEf+B8/gBzSo3t</latexit> x s <latexit sha1_base64="v7xFkNYPbfbModP8QyvO1wHTzxo=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120i7dbMLuRiyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEsG1cd1vZ2V1bX1js7BV3N7Z3dsvHRw2dZwqhg0Wi1i1A6pRcIkNw43AdqKQRoHAVjC6mfqtR1Sax/LBjBP0IzqQPOSMGivdP/WwVyq7FXcGsky8nJQhR71X+ur2Y5ZGKA0TVOuO5ybGz6gynAmcFLupxoSyER1gx1JJI9R+Njt1Qk6t0idhrGxJQ2bq74mMRlqPo8B2RtQM9aI3Ff/zOqkJr/yMyyQ1KNl8UZgKYmIy/Zv0uUJmxNgSyhS3txI2pIoyY9Mp2hC8xZeXSfO84lUr1buLcu06j6MAx3ACZ+DBJdTgFurQAAYDeIZXeHOE8+K8Ox/z1hUnnzmCP3A+fwBeEo3f</latexit> x e <latexit sha1_base64="iANey9rYlpgkCfH1CXEhuEbhg2o=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSIIQklEqseiF48V7Ae0oWy2k3bpZhN2N2IJ/RFePCji1d/jzX/jts1BWx8MPN6bYWZekAiujet+Oyura+sbm4Wt4vbO7t5+6eCwqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLR7dRvPaLSPJYPZpygH9GB5CFn1Fip9dTL9Lk36ZXKbsWdgSwTLydlyFHvlb66/ZilEUrDBNW647mJ8TOqDGcCJ8VuqjGhbEQH2LFU0gi1n83OnZBTq/RJGCtb0pCZ+nsio5HW4yiwnRE1Q73oTcX/vE5qwms/4zJJDUo2XxSmgpiYTH8nfa6QGTG2hDLF7a2EDamizNiEijYEb/HlZdK8qHjVSvX+sly7yeMowDGcwBl4cAU1uIM6NIDBCJ7hFd6cxHlx3p2PeeuKk88cwR84nz8Rm49p</latexit> x s+1 <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · Selected Spans Stage 2: Test-Time Training on Selected Spans <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="7fJgorBTuC0IU+mE+H26FyrGvEQ=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qlneqWyW3FnIMvEy0kZctR7pa9uP2ZpxBUySY3peG6CfkY1Cib5pNhNDU8oG9EB71iqaMSNn81OnZBTq/RJGGtbCslM/T2R0ciYcRTYzoji0Cx6U/E/r5NieOVnQiUpcsXmi8JUEozJ9G/SF5ozlGNLKNPC3krYkGrK0KZTtCF4iy8vk+Z5xatWqncX5dp1HkcBjuEEzsCDS6jBLdShAQwG8Ayv8OZI58V5dz7mrStOPnMEf+B8/gBzSo3t</latexit> x s <latexit sha1_base64="v7xFkNYPbfbModP8QyvO1wHTzxo=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120i7dbMLuRiyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEsG1cd1vZ2V1bX1js7BV3N7Z3dsvHRw2dZwqhg0Wi1i1A6pRcIkNw43AdqKQRoHAVjC6mfqtR1Sax/LBjBP0IzqQPOSMGivdP/WwVyq7FXcGsky8nJQhR71X+ur2Y5ZGKA0TVOuO5ybGz6gynAmcFLupxoSyER1gx1JJI9R+Njt1Qk6t0idhrGxJQ2bq74mMRlqPo8B2RtQM9aI3Ff/zOqkJr/yMyyQ1KNl8UZgKYmIy/Zv0uUJmxNgSyhS3txI2pIoyY9Mp2hC8xZeXSfO84lUr1buLcu06j6MAx3ACZ+DBJdTgFurQAAYDeIZXeHOE8+K8Ox/z1hUnnzmCP3A+fwBeEo3f</latexit> x e <latexit sha1_base64="iANey9rYlpgkCfH1CXEhuEbhg2o=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSIIQklEqseiF48V7Ae0oWy2k3bpZhN2N2IJ/RFePCji1d/jzX/jts1BWx8MPN6bYWZekAiujet+Oyura+sbm4Wt4vbO7t5+6eCwqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLR7dRvPaLSPJYPZpygH9GB5CFn1Fip9dTL9Lk36ZXKbsWdgSwTLydlyFHvlb66/ZilEUrDBNW647mJ8TOqDGcCJ8VuqjGhbEQH2LFU0gi1n83OnZBTq/RJGCtb0pCZ+nsio5HW4yiwnRE1Q73oTcX/vE5qwms/4zJJDUo2XxSmgpiYTH8nfa6QGTG2hDLF7a2EDamizNiEijYEb/HlZdK8qHjVSvX+sly7yeMowDGcwBl4cAU1uIM6NIDBCJ7hFd6cxHlx3p2PeeuKk88cwR84nz8Rm49p</latexit> x s+1 <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="0qSiA+TXOQEOvRSqdJWQbG5jZ2U=">AAAB6nicbVDLTgJBEOzFF+IL9ehlIjHxRHaJQY9ELx4xyiOBDZkdemHC7OxmZtZICJ/gxYPGePWLvPk3DrAHBSvppFLVne6uIBFcG9f9dnJr6xubW/ntws7u3v5B8fCoqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLRzcxvPaLSPJYPZpygH9GB5CFn1Fjp/qlX6RVLbtmdg6wSLyMlyFDvFb+6/ZilEUrDBNW647mJ8SdUGc4ETgvdVGNC2YgOsGOppBFqfzI/dUrOrNInYaxsSUPm6u+JCY20HkeB7YyoGeplbyb+53VSE175Ey6T1KBki0VhKoiJyexv0ucKmRFjSyhT3N5K2JAqyoxNp2BD8JZfXiXNStmrlqt3F6XadRZHHk7gFM7Bg0uowS3UoQEMBvAMr/DmCOfFeXc+Fq05J5s5hj9wPn8AEMaNrA==</latexit> x 2 <latexit sha1_base64="P4GArwQ3oaXEfJT4toixLcIzE8A=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qnn9Uplt+LOQJaJl5My5Kj3Sl/dfszSiCtkkhrT8dwE/YxqFEzySbGbGp5QNqID3rFU0YgbP5udOiGnVumTMNa2FJKZ+nsio5Ex4yiwnRHFoVn0puJ/XifF8MrPhEpS5IrNF4WpJBiT6d+kLzRnKMeWUKaFvZWwIdWUoU2naEPwFl9eJs3ziletVO8uyrXrPI4CHMMJnIEHl1CDW6hDAxgM4Ble4c2Rzovz7nzMW1ecfOYI/sD5/AEPQo2r</latexit> x 1 <latexit sha1_base64="Ly9Q9FAkfVUNNdexplHG7d2WuvE=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSJ4sSQi1WPRi8cK9gPaUDbbSbt0swm7G7GE/ggvHhTx6u/x5r9x2+agrQ8GHu/NMDMvSATXxnW/nZXVtfWNzcJWcXtnd2+/dHDY1HGqGDZYLGLVDqhGwSU2DDcC24lCGgUCW8Hoduq3HlFpHssHM07Qj+hA8pAzaqzUeupl+tyb9Eplt+LOQJaJl5My5Kj3Sl/dfszSCKVhgmrd8dzE+BlVhjOBk2I31ZhQNqID7FgqaYTaz2bnTsipVfokjJUtachM/T2R0UjrcRTYzoiaoV70puJ/Xic14bWfcZmkBiWbLwpTQUxMpr+TPlfIjBhbQpni9lbChlRRZmxCRRuCt/jyMmleVLxqpXp/Wa7d5HEU4BhO4Aw8uIIa3EEdGsBgBM/wCm9O4rw4787HvHXFyWeO4A+czx8Up49r</latexit> x s→1 Selected SpanPrefix Next-Token Prediction Training Input (previous token) Target (selected span) <latexit sha1_base64="Ly9Q9FAkfVUNNdexplHG7d2WuvE=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSJ4sSQi1WPRi8cK9gPaUDbbSbt0swm7G7GE/ggvHhTx6u/x5r9x2+agrQ8GHu/NMDMvSATXxnW/nZXVtfWNzcJWcXtnd2+/dHDY1HGqGDZYLGLVDqhGwSU2DDcC24lCGgUCW8Hoduq3HlFpHssHM07Qj+hA8pAzaqzUeupl+tyb9Eplt+LOQJaJl5My5Kj3Sl/dfszSCKVhgmrd8dzE+BlVhjOBk2I31ZhQNqID7FgqaYTaz2bnTsipVfokjJUtachM/T2R0UjrcRTYzoiaoV70puJ/Xic14bWfcZmkBiWbLwpTQUxMpr+TPlfIjBhbQpni9lbChlRRZmxCRRuCt/jyMmleVLxqpXp/Wa7d5HEU4BhO4Aw8uIIa3EEdGsBgBM/wCm9O4rw4787HvHXFyWeO4A+czx8Up49r</latexit> x s→1 <latexit sha1_base64="7fJgorBTuC0IU+mE+H26FyrGvEQ=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qlneqWyW3FnIMvEy0kZctR7pa9uP2ZpxBUySY3peG6CfkY1Cib5pNhNDU8oG9EB71iqaMSNn81OnZBTq/RJGGtbCslM/T2R0ciYcRTYzoji0Cx6U/E/r5NieOVnQiUpcsXmi8JUEozJ9G/SF5ozlGNLKNPC3krYkGrK0KZTtCF4iy8vk+Z5xatWqncX5dp1HkcBjuEEzsCDS6jBLdShAQwG8Ayv8OZI58V5dz7mrStOPnMEf+B8/gBzSo3t</latexit> x s <latexit sha1_base64="VfG5dHVvf4FXm6iZfsTexVMdeUI=">AAAB73icbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04kkq2A9oQ9lsN+3SzSbuToQS+ie8eFDEq3/Hm//GbZuDtj4YeLw3w8y8IJHCoOt+Oyura+sbm4Wt4vbO7t5+6eCwaeJUM95gsYx1O6CGS6F4AwVK3k40p1EgeSsY3Uz91hPXRsTqAccJ9yM6UCIUjKKV2l0UETfkrlcquxV3BrJMvJyUIUe9V/rq9mOWRlwhk9SYjucm6GdUo2CST4rd1PCEshEd8I6lito1fja7d0JOrdInYaxtKSQz9fdERiNjxlFgOyOKQ7PoTcX/vE6K4ZWfCZWkyBWbLwpTSTAm0+dJX2jOUI4toUwLeythQ6opQxtR0YbgLb68TJrnFa9aqd5flGvXeRwFOIYTOAMPLqEGt1CHBjCQ8Ayv8OY8Oi/Ou/Mxb11x8pkj+APn8wevfo/B</latexit> →N steps <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · <latexit sha1_base64="7fJgorBTuC0IU+mE+H26FyrGvEQ=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qlneqWyW3FnIMvEy0kZctR7pa9uP2ZpxBUySY3peG6CfkY1Cib5pNhNDU8oG9EB71iqaMSNn81OnZBTq/RJGGtbCslM/T2R0ciYcRTYzoji0Cx6U/E/r5NieOVnQiUpcsXmi8JUEozJ9G/SF5ozlGNLKNPC3krYkGrK0KZTtCF4iy8vk+Z5xatWqncX5dp1HkcBjuEEzsCDS6jBLdShAQwG8Ayv8OZI58V5dz7mrStOPnMEf+B8/gBzSo3t</latexit> x s <latexit sha1_base64="iANey9rYlpgkCfH1CXEhuEbhg2o=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSIIQklEqseiF48V7Ae0oWy2k3bpZhN2N2IJ/RFePCji1d/jzX/jts1BWx8MPN6bYWZekAiujet+Oyura+sbm4Wt4vbO7t5+6eCwqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLR7dRvPaLSPJYPZpygH9GB5CFn1Fip9dTL9Lk36ZXKbsWdgSwTLydlyFHvlb66/ZilEUrDBNW647mJ8TOqDGcCJ8VuqjGhbEQH2LFU0gi1n83OnZBTq/RJGCtb0pCZ+nsio5HW4yiwnRE1Q73oTcX/vE5qwms/4zJJDUo2XxSmgpiYTH8nfa6QGTG2hDLF7a2EDamizNiEijYEb/HlZdK8qHjVSvX+sly7yeMowDGcwBl4cAU1uIM6NIDBCJ7hFd6cxHlx3p2PeeuKk88cwR84nz8Rm49p</latexit> x s+1 <latexit sha1_base64="v7xFkNYPbfbModP8QyvO1wHTzxo=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120i7dbMLuRiyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEsG1cd1vZ2V1bX1js7BV3N7Z3dsvHRw2dZwqhg0Wi1i1A6pRcIkNw43AdqKQRoHAVjC6mfqtR1Sax/LBjBP0IzqQPOSMGivdP/WwVyq7FXcGsky8nJQhR71X+ur2Y5ZGKA0TVOuO5ybGz6gynAmcFLupxoSyER1gx1JJI9R+Njt1Qk6t0idhrGxJQ2bq74mMRlqPo8B2RtQM9aI3Ff/zOqkJr/yMyyQ1KNl8UZgKYmIy/Zv0uUJmxNgSyhS3txI2pIoyY9Mp2hC8xZeXSfO84lUr1buLcu06j6MAx3ACZ+DBJdTgFurQAAYDeIZXeHOE8+K8Ox/z1hUnnzmCP3A+fwBeEo3f</latexit> x e <latexit sha1_base64="Dknxhvd5O3dvukcomWGfhj7K5fA=">AAAB7nicbVBNS8NAEJ34WetX1aOXxSJ4sSQi1WPRi8cK9gPaUDbbSbt0swm7G7GE/ggvHhTx6u/x5r9x2+agrQ8GHu/NMDMvSATXxnW/nZXVtfWNzcJWcXtnd2+/dHDY1HGqGDZYLGLVDqhGwSU2DDcC24lCGgUCW8Hoduq3HlFpHssHM07Qj+hA8pAzaqzUeupleO5NeqWyW3FnIMvEy0kZctR7pa9uP2ZphNIwQbXueG5i/Iwqw5nASbGbakwoG9EBdiyVNELtZ7NzJ+TUKn0SxsqWNGSm/p7IaKT1OApsZ0TNUC96U/E/r5Oa8NrPuExSg5LNF4WpICYm099JnytkRowtoUxxeythQ6ooMzahog3BW3x5mTQvKl61Ur2/LNdu8jgKcAwncAYeXEEN7qAODWAwgmd4hTcncV6cd+dj3rri5DNH8AfO5w//No9d</latexit> x e→1 Trained LLM Inference <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · Long Context Tokens <latexit sha1_base64="0qSiA+TXOQEOvRSqdJWQbG5jZ2U=">AAAB6nicbVDLTgJBEOzFF+IL9ehlIjHxRHaJQY9ELx4xyiOBDZkdemHC7OxmZtZICJ/gxYPGePWLvPk3DrAHBSvppFLVne6uIBFcG9f9dnJr6xubW/ntws7u3v5B8fCoqeNUMWywWMSqHVCNgktsGG4EthOFNAoEtoLRzcxvPaLSPJYPZpygH9GB5CFn1Fjp/qlX6RVLbtmdg6wSLyMlyFDvFb+6/ZilEUrDBNW647mJ8SdUGc4ETgvdVGNC2YgOsGOppBFqfzI/dUrOrNInYaxsSUPm6u+JCY20HkeB7YyoGeplbyb+53VSE175Ey6T1KBki0VhKoiJyexv0ucKmRFjSyhT3N5K2JAqyoxNp2BD8JZfXiXNStmrlqt3F6XadRZHHk7gFM7Bg0uowS3UoQEMBvAMr/DmCOfFeXc+Fq05J5s5hj9wPn8AEMaNrA==</latexit> x 2 <latexit sha1_base64="P4GArwQ3oaXEfJT4toixLcIzE8A=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120y7dbMLuRCyhP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEikMuu63s7K6tr6xWdgqbu/s7u2XDg6bJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LRzdRvPXJtRKwecJxwP6IDJULBKFrp/qnn9Uplt+LOQJaJl5My5Kj3Sl/dfszSiCtkkhrT8dwE/YxqFEzySbGbGp5QNqID3rFU0YgbP5udOiGnVumTMNa2FJKZ+nsio5Ex4yiwnRHFoVn0puJ/XifF8MrPhEpS5IrNF4WpJBiT6d+kLzRnKMeWUKaFvZWwIdWUoU2naEPwFl9eJs3ziletVO8uyrXrPI4CHMMJnIEHl1CDW6hDAxgM4Ble4c2Rzovz7nzMW1ecfOYI/sD5/AEPQo2r</latexit> x 1 <latexit sha1_base64="kfyhpuCjonc2/DMNNayB027cIxc=">AAAB6nicbVDLSgNBEOyNrxhfUY9eBoPgKeyKRI9BLx4j5gXJEmYns8mQ2dllplcMIZ/gxYMiXv0ib/6Nk2QPmljQUFR1090VJFIYdN1vJ7e2vrG5ld8u7Ozu7R8UD4+aJk414w0Wy1i3A2q4FIo3UKDk7URzGgWSt4LR7cxvPXJtRKzqOE64H9GBEqFgFK308NSr94olt+zOQVaJl5ESZKj1il/dfszSiCtkkhrT8dwE/QnVKJjk00I3NTyhbEQHvGOpohE3/mR+6pScWaVPwljbUkjm6u+JCY2MGUeB7YwoDs2yNxP/8zophtf+RKgkRa7YYlGYSoIxmf1N+kJzhnJsCWVa2FsJG1JNGdp0CjYEb/nlVdK8KHuVcuX+slS9yeLIwwmcwjl4cAVVuIMaNIDBAJ7hFd4c6bw4787HojXnZDPH8AfO5w9ETo3O</latexit> x T <latexit sha1_base64="HAVxy/77DfEejtoQbW/WOeu2/bM=">AAAB6nicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rGi/YA2lM120i7dbOLuRiihP8GLB0W8+ou8+W/ctjlo64OBx3szzMwLEsG1cd1vZ2V1bX1js7BV3N7Z3dsvHRw2dZwqhg0Wi1i1A6pRcIkNw43AdqKQRoHAVjC6mfqtJ1Sax/LBjBP0IzqQPOSMGivdP/a8XqnsVtwZyDLxclKGHPVe6avbj1kaoTRMUK07npsYP6PKcCZwUuymGhPKRnSAHUsljVD72ezUCTm1Sp+EsbIlDZmpvycyGmk9jgLbGVEz1IveVPzP66QmvPIzLpPUoGTzRWEqiInJ9G/S5wqZEWNLKFPc3krYkCrKjE2naEPwFl9eJs3ziletVO8uyrXrPI4CHMMJnIEHl1CDW6hDAxgM4Ble4c0Rzovz7nzMW1ecfOYI/sD5/AEEmI2k</latexit> q 1 <latexit sha1_base64="qb/tfXDey1ZHG0Gq1YoCB5K5dGA=">AAAB6nicbVDLTgJBEOzFF+IL9ehlIjHxRHaJQY9ELx4xyiOBDZkdZmHC7Ow602tCCJ/gxYPGePWLvPk3DrAHBSvppFLVne6uIJHCoOt+O7m19Y3Nrfx2YWd3b/+geHjUNHGqGW+wWMa6HVDDpVC8gQIlbyea0yiQvBWMbmZ+64lrI2L1gOOE+xEdKBEKRtFK94+9Sq9YcsvuHGSVeBkpQYZ6r/jV7ccsjbhCJqkxHc9N0J9QjYJJPi10U8MTykZ0wDuWKhpx40/mp07JmVX6JIy1LYVkrv6emNDImHEU2M6I4tAsezPxP6+TYnjlT4RKUuSKLRaFqSQYk9nfpC80ZyjHllCmhb2VsCHVlKFNp2BD8JZfXiXNStmrlqt3F6XadRZHHk7gFM7Bg0uowS3UoQEMBvAMr/DmSOfFeXc+Fq05J5s5hj9wPn8ABhyNpQ==</latexit> q 2 <latexit sha1_base64="sT9mS3Xirev4s8d64IIm1mu0U1I=">AAAuf3icxVpbc9vGFaaTtE3Vm9M+9mWnHoVkRXIAkLo9qJPGaaeeOhrXE9nOiBQCgiCJiARoAhRJQfsv+9K/kqees7juYgGCsmeiiSUAe853LntuC2S4mNmeryj/e/LJp5/94pe/+vzXB7/57e9+/4enX/zxjeeulqZ1Zbozd/luaHjWzHasK9/2Z9a7xdIy5sOZ9XZ4+xzX395ZS892ne/87cIazI2JY49t0/Dhkf5Fa37YH7qzkbedw5/gjpILcn2nqy1yp2st0h+5vofXlwPStx3Snxv+dDgMXtObywMJJ/ApLYK/c/Q90n+/MkYky7RmTCGHnCknpG+CSnmQvrea64F9oYJioK691m14qgK7b238AHkWS3e0Mn3a4BFbAlgT0SrJBDetdZUcoavWugYXjNBjT5y17sj1Ojh84PEfGN37pR/slHtHaUosQN9oFMw1XU80UGLfeGmYu6WtKQ1yuj7wFA80xcvbSoNCXWVWrNkKLfOQBCheQ0wWYqKKcu61yP3zek+loQoY/KYxC96wpSCM34nrjjBS2c3YsmbJzYvkyjNSEkhvL7mZGncW7VMxlfQsdiZ70/8G/X4RE9Mhk72VmF4kHBxTCQfalOFJ2Up4mO08U1yTipmYj3JMyMeYRpDnX15E/vzhhxcE7SfouXqdhgSahADUT9a7/DoKJKhqRCCqBRIpclyrSXUssRzEh9RJLS3zEyiTUCcuVWNbJUkQqtOSSU0TQqNBl+4L0M0AqDsAtOoAZc5Uey3SPWkRbbcrz+Df8RnwHJf78bRFTjTmQq2AsBfJ7wLoeYugCiXRODR86OQ00RnUABFqGUuUxYwBLUTVUaez87Isdt1ZxIQeYaaCciUca9uPGDSMmWOwuKIpXH3ZbQhXWXabwAXzTv25ghJSs9Xv/kkb69YoHAGsie1AMfEsD0ontvn+zJ3ogarQuMi6K8ePGL6UPPsbUQhiK8mq60+t5dr2LCzGljNK4MPlb1B8k58agksaD0RzCPDGmo1I91uxlwvUrZUwZk19hthVlzSIAtP+0t7MvWbqc3tECGym3dE37fnlkc4SQUJmt2PXKJmvI9SlM55tzDTKwDFURViKecfgpUq9dE0elwRUzqqwgaYQuaujLkXe/Qc+OSCpTOuSLYWyO4o7PSi8fcW+Rrp4aoJv+Aan67xiABDL0Y5u2uyPxquhg559e0LGsekFhXoDB8NOLYcaf9Zlja8g3B8FlInQmIpc2MTxnxlgS3MzjgVFlECJaJRJkvCdT09k4RjGcX1epM1moUeoCPhlHW9nAwHAWzhWe/4XG0pnV5XUXvHcHGuKKdnJ4BnzBZTgxKak3VDgj1BAGWXfjd7Y2anz5eRV+O7OMzhcDm2lpZjYknHkghHzZEB05GEdmSzk6e/5Ug1GanpOh4cdgEXiYuKt+/K59TonFrWhbbyYXU3pz81fPnYupuXa2UC+y5exzYlA6lazdrc1JxwV1e5rXa6MEdoHZwKlE4XBpBepwvdvt3taBWQUgPaCsNQO+eniISDQ7eDk1A7BN+F5DoTYk7dssjwFrZjmNOQAmqBzh3P+tB2fUyaPr4D0ANobWaTT5tvYHUh52qYD6znQV0pJWHoC98looqm/CQPuRvXkGLMpE5Ym0WuIZQAJ/XLrKOdD5n68EDlYPUyNLC/EK7YK/XQBLV8UxhwwZ7s9vrFLgA2CVX3HGmTUOrH9F/aaQqNiQav5cqiD2YrMsyezI29tj17FmIzK6ra3itwPmSzkg78KN3bic957Ms4EIpAq7l+exHxb5ph1oV1yppQ1pKT9I2s2GcWiIXcQKM9PRZetUTOuwka/pGKtBPL90gBRRMDsA93pO8Yw5lRtA9pD+0P7Ul/Ca3Wl9JeFAoqUPOvBePnGk/weQ/jO4CDw2VaqZhXn7t39Pr19btBi7y+/n4wQBei7/UAH9L05vuBqEda1d+vLMtBY40F7M0mJzua3925gVTtgvVw9ahg9dZ2JqjBmr2eMZaTMPjXdXGn8wMzggxlcg2ZOFPmvDrGJ/SAnbKZaz9IerMgQ+r515T7AT/kuAHzQdjUTTiPdI5xAjjBX+qAp4i+BXQ0Nm3gr3OgGLKjzfHB4b3wWn4tS/kNahiynJ0eHBppDb0vqIxI2EwjDLJXwbOF4UyzLNZNcI8egb/te/BXeH8U38enkdfWyysa8cFONu7BiiZv5VDqlw0eUjaZDy4QBY6uUMFFb6n4fQSo1PiMjgxElCZjEFS4khFpKaoqok5p4lfSEPUTd4NXR/DGPRV29YpHmAqqbmn6EtVzxz44mf+ccw9hDunEqLCMX2emp3s4/MWV/Ef2asXBV33R2o94Ioy8zzExqnLGqPoPuPNTXGCe/wObO5wiAt4SoRxsWe+O+BtKmar30bNDzG2VL/N+SSc9CPl1eRK+TgKZWWbmUeck0zRRf/OOkhoNFwal/XFdYWWuOa7YTXLYBjgagMo+LGg00a4MJa+bcyyEDTzVBD0CLZt0hmKluUL9+V89yVq7pmPuXwOX1MKmZklWVFdzRf2qRhewgmkPIkFfIH5Q6TFmU/wk2WGZmfdHtXjAtu4jBrZqCkW2o1YHjeYDyxV3MVs5eXXtMxa+C26iPKymVcq0YQpwi+OxB4QZWbSCeKxmbNbn5I8VxstOBIprbyOfjpYiw7Vrco9qS4A15uil+vRS6XKbbhe1ocbwMLqmazTy3azJd3HVrSDklXYu4FkOxSx0G5jJ+aSNH8qzmVqyBq5n+wcGbjwj13BAonPvw0rAYtQgU3ogFhI5ikLZDjYPJBNHPIRURTXGVpHpG3F24eL7SzzRfwKyMe8drIofhkKi1lpKG7TkM55WfLyQepmTkVZnjGCx2SaVjXTOBX6d6bl+Pj22Lkgk7wRKbgkOCtYoz3Smm5la7T9rekWWKOVW/O4nVEqWrL3rihSK4QdyQKN80CZusQyQO6Uoq3PyMnPni+/pVRSDpKDmRO/YYrys2B8TRwtGcjCVBbHMr6blIkUJt4gqg2ZfreLMxk8uOISiU7fJeUP+ThW4NskcF3hKJj1ZIrfy+LLRrc4OGRzpiSkd6LswtDKMdQqGN1yDK0KRq8cA0WQksH7Q7zdawpf1HaB8dVlXDBh8d1EH1cesMaShC0asOxqwu3Kwu09hE8ywsNvcIYzlTdTfVJZg8keGpgZDWRe6bvikSHR+YgQ6S7mWcxEWV64W833bmXL3T0sn8otdwUz0m3J62bmBPh69yAc2/Snz5SOwn5I/kKNLp7Vop9X+tOfoAeZqzk2rJnhedeqsvAHAeacObPoQX/lWQvDvDUm1jVcOgaMnYOAffim5BCejMjYXcI/xyfsaZYjMOYeKg2UmKWeuIYPZWvXK398NghsZ7HCT9ahoPFqRnyX4P9kTUb20jL92RYuDHNpg67EnBpQZHxryUsZDuecEQEK81135tEDcJYquiZ/8UbrqCedk//0n31deS2z2t/rv2l1qiptdPaV7V/1V7Vrmpm67+tn9qftj/rPOnUO52OEpJ+8iTi+VON++mc/x8j0bQT</latexit> · Question Tokens Answer y <latexit sha1_base64="Wc+SuwGu1gWPWB5JAP5olpWFkag=">AAAB6HicbVBNS8NAEJ34WetX1aOXxSJ4KolI9Vj04rEF+wFtKJvtpF272YTdjRBKf4EXD4p49Sd589+4bXPQ1gcDj/dmmJkXJIJr47rfztr6xubWdmGnuLu3f3BYOjpu6ThVDJssFrHqBFSj4BKbhhuBnUQhjQKB7WB8N/PbT6g0j+WDyRL0IzqUPOSMGis1sn6p7FbcOcgq8XJShhz1fumrN4hZGqE0TFCtu56bGH9CleFM4LTYSzUmlI3pELuWShqh9ifzQ6fk3CoDEsbKljRkrv6emNBI6ywKbGdEzUgvezPxP6+bmvDGn3CZpAYlWywKU0FMTGZfkwFXyIzILKFMcXsrYSOqKDM2m6INwVt+eZW0LitetVJtXJVrt3kcBTiFM7gAD66hBvdQhyYwQHiGV3hzHp0X5935WLSuOfnMCfyB8/kD60+NCA==</latexit> y q M q M Figure 1 Overview of Self-Guided T. In Stage 1, the base LLM reads the long context and question, and identifies question-relevant spans from the context. In Stage 2, these selected spans are used for T. At inference time, the adapted model generates the answer conditioning on the original full context and the question. In contrast, when the training spans are oracle spans annotated by GPT-5.5 with access to the ground-truth answer, the same T procedure achieves 45.9% accuracy. Notably, we explicitly control the length of oracle spans to be comparable to that of the random spans, therefore, the number of training tokens is not the factor affecting the performance. This gap isolates the role of the training data quality: T can help, but only when the tokens used contain useful evidence. This motivates our core view: the central bottleneck in long-context T is not only how to adapt the model, but also what to adapt on. High-quality spans provide a much stronger training signal, however, relying on an external oracle is not a practical solution. We therefore ask whether the model can identify the effective test-time training tokens by itself. 2.2 Self-Guided T The preliminary analysis suggests that test-time training is effective only when the training tokens provide useful information for the current test instance. Based on this observation, we introduce Self-Guided T (S-T), in which the model first identifies question-relevant evidence from the context and then adapts itself on the selected evidence. Specifically, given a contextx= (x 1 ,...,x T ), a questionq, and a base model with parameters θ, S-T consists of two stages: Stage 1: Model-guided span selection. We first ask the model to identify the parts of the context that are most relevant to answeringq. Concretely, the model reads the full context and question and returns a set of verbatim supporting spans, S(x,q) = x s j :e j M j=1 , where each interval [s j ,e j ] corresponds to a contiguous span copied from the original context. This selection step relies on the model’s own relevance judgment to construct instance-specific training data. The purpose of this stage is not to replace the original context at generation time, but to identify the subset of tokens that provides the most useful adaptation signal, which may otherwise be buried among a large amount of irrelevant information in the full context. Stage 2: Test-time training on selected spans. Starting from a fresh copy of the base modelθ ′ ← θ, we perform next-token prediction on the selected spans. For a selected span x s j :e j , the training objective is L T (θ ′ ) =− e j X i=s j logp θ ′ (x i | x <i ). Across adaptation steps, we cycle through the valid spans inS(x,q) and updateθ ′ using the training objective above. The model is encouraged to internalize information that is likely to be useful for answering the current question identified by itself, rather than arbitrary content from the long context. 3 Algorithm 1 Self-Guided T Require: Model θ, context x 1:T , question q, steps N, span length k, learning rate η 1: Initialize a fresh model θ ′ from θ 2: S ← spans in x 1:T annotated by θ relevant to q 3: if S =∅ then 4: S ← random spans sampled from x 1:T 5: end if 6: for n = 1,...,N do 7: Choose span x s j :e j from S 8: L T ←− P e j i=s j logp θ ′ (x i | x <i ) 9: Update θ ′ ← θ ′ − η∇L T 10: end for 11: return answer y ∼ p θ ′ (·| x 1:T ,q) After adaptation, the updated model generates the answer conditioned on the original full context and question: y ∼ p θ ′ (·| x 1:T ,q). The full context remains available during generation, so span selection determines only the test-time training data and does not remove potentially useful information from the final input. Once the instance is completed, θ ′ is discarded and the next instance begins from the original parametersθ. A per-instance loop is described in Algorithm 1. 3 Experimental Results 3.1 Setup Models and benchmarks. We evaluate two base models: Qwen3-4B-Thinking-2507 (Qwen Team, 2025) and Llama-3.1-8B-Instruct (Team, 2024). We conduct experiments on two challenging long-context benchmarks. LongBench-v2 (Bai et al., 2025) is a four-way multiple-choice benchmark covering diverse long-context reasoning tasks and is evaluated using answer accuracy. LongBench-Pro (Chen et al., 2026b) evaluates a broader set of long-context capabilities and has English and Chinese subsets, we use its English subset as our evaluation set and apply its official evaluation pipeline for scoring. We use the Qwen3 tokenizer to measure context length and keep examples whose contexts contain at most 128k tokens. Compared methods. We compare the following settings: •Base Model. The base model directly generates an answer conditioned on the full context and question, without any parameter updates. •LongLLMLingua. LongLLMLingua (Jiang et al., 2024b) is a prompt compression method. The base model first compresses the full context into a shorter one conditioned on the question, and then answers the question using the compressed context. We set the compressed-context budget to be 4,096 tokens. •qTTT. qTTT (Bansal et al., 2026) is an efficiency-oriented T method that adapts on uniformly sampled random spans. It first runs a single forward pass over the full context to build the KV cache, then keeps the cache frozen and updates only the query-projection parameters, avoiding recomputation of the full-context KV at every adaptation step. • QRHead Span T. Following QRHead (Zhang et al., 2025b), we identify query-relevant attention heads on a retrieval set BEIR (Thakur et al., 2021) by scoring each head’s query-to-context attention as a retriever and the 16 highest-scoring heads are kept as QRHeads. At test time, we run a single forward pass over the full context to obtain QRHead attention scores, aggregate them over each 512-token candidate span to obtain span-level scores, and select the 8 highest-scoring spans for T. 4 Table 2 Results on LongBench-v2 and LongBench-Pro across different context-length buckets. The best result for each model and evaluation setting is shown in bold. † QRHead Span T is not directly comparable with the other T methods as it requires additional information to identify retrieval heads. ModelMethod LongBench-v2LongBench-Pro < 64k64k–128k< 64k64k–128k Qwen3-4B- Thinking-2507 Base Model46.730.755.141.6 LongLLMLingua41.831.735.030.3 qTTT44.734.056.641.5 QRHead Span T † 47.232.156.740.8 Random Span T43.634.255.041.0 Full Context T45.132.655.840.4 Self-Guided T47.735.356.242.0 Llama-3.1- 8B-Instruct Base Model36.926.328.219.4 LongLLMLingua34.126.925.221.0 qTTT35.727.529.719.3 QRHead Span T † 35.727.529.420.4 Random Span T36.026.728.720.4 Full Context T35.227.729.419.8 Self-Guided T38.428.229.921.7 •Random Span T. We randomly sample 8 spans from the context, each containing 512 tokens, and use them for T. Random Span T differs from qTTT in updating with a non-frozen KV cache. •Full Context T. We partition the full context intoNcontiguous chunks, whereNis the number of adaptation steps. At each step, we perform one step T update on one chunk. •Self-Guided T (ours). We ask the model to identify at most 8 spans in the context that are relevant to the question and then perform T on the selected spans. If the model fails to output valid spans, it falls back to using uniformly sampled spans. We report fallback rates in Appendix B. For all T methods, the final answer is generated conditioned on the full context rather than the selected spans. We use LoRA for parameter-efficient test-time adaptation and perform 16 gradient-update steps for each test instance. More details can be found in Appendix A. Evaluation. For each test instance, we sample 4 responses and evaluate each response using the corresponding benchmark evaluator. We report the mean scores for the four samples. We sample responses with a temperature of 0.6 and a top-pof 0.95. The maximum generation length is set to 32,768 tokens for Qwen3-4B-Thinking-2507 and 10,240 tokens for Llama-3.1-8B-Instruct. Prompt templates can be found in Appendix C. 3.2 Results Table 2 summarizes the main results. We highlight three observations. First, S-T consistently improves over the base model across models, benchmarks, and length buckets. Second, S-T consistently outperforms or is comparable to all the other T methods, showing that model-annotated training spans provide a more reliable adaptation signal than uniformly sampled spans or full context. Third, the gains are especially pronounced in the longer-context buckets, where irrelevant context is more abundant and training-token selection becomes more important. LongBench-v2. Using Qwen3-4B-Thinking-2507 as the base model, Random Span T degrades the<64k bucket, reducing accuracy from 46.7 to 43.6, whereas S-T improves it to 47.7. In the 64k–128k bucket, all T baselines help, but S-T leads to the strongest score, reaching 35.3. QRHead Span T is competitive 5 in the shorter bucket, but drops behind in the longer bucket, suggesting that attention-based span scores are less stable as the context grows. Other baselines are less consistent: LongLLMLingua often underperforms the base model, and Full Context T remains below S-T in every setting. The trend also transfers to Llama-3.1-8B-Instruct. S-T gives the best LongBench-v2 scores in both buckets, improving the base model from 36.9 to 38.4 and from 26.3 to 28.2, respectively. This indicates that the gain from S-T is model-agnostic. LongBench-Pro. With Qwen3-4B-Thinking-2507, S-T improves over the base model in both length buckets. In the shorter bucket, qTTT and QRHead Span T are slightly higher than S-T. In the longer bucket, however, S-T is the strongest method, reaching 42.0 and outperforming all T baselines. This mirrors the LongBench-v2 trend: model-annotated spans become more valuable when the context is longer and noisier. With Llama-3.1-8B-Instruct, S-T again gives the best LongBench-Pro scores in both buckets, improving the base model from 28.2 to 29.9 and from 19.4 to 21.7, respectively. Overall, these results support our main hypothesis that training-token quality is a central bottleneck for long-context T. 4 Analysis We further analyze why and when S-T works. We ask three questions: (1) whether question-conditioned model annotation is more effective than annotation-free intrinsic span scores, (2) how adaptation on selected spans changes the model’s attention to the relevant evidence, and (3) how the end-to-end overhead of S-T scales with context length. 4.1 Span selection strategies The main results show that span selection matters. We next ask whether explicit model annotation is necessary, or whether simpler annotation-free signals can select useful training spans. A natural alternative is to use intrinsic model statistics, selecting spans that are difficult to predict or induce high uncertainty in the next-token distribution. We compare against two intrinsic selectors. The perplexity selector ranks each 512-token window by mean negative log-likelihood, while the entropy selector ranks each window by mean predictive entropy. Each method then performs T on the top 8 highest-scoring spans. We keep the training configuration fixed across all methods, changing only how the training spans are selected. Span selector<64k 64–128k Model annotation 47.735.3 Perplexity score46.731.9 Entropy score45.133.0 Table 3 Model-annotated spans outperform intrinsic metric-selected spans on LongBench-v2 with Qwen3-4B-Thinking- 2507. Table 3 shows that model-annotated spans perform best in both length buckets, indicating that model intrinsic metrics are not the best choice for selecting useful T data. The gap is small below 64k for perplexity-selected spans but becomes much larger in the 64k–128k bucket. This is the regime where the context contains more distractors and where selecting question-relevant evidence becomes most important. These results suggest that useful T spans are not simply the spans that are surprising or uncertain under the language model. High-perplexity or high-entropy text may be difficult to predict for many reasons unrelated to the question, such as formatting, rare entities, or local distribution shift. In contrast, model annotation conditions span selection on the question, allowing it to better target evidence that can improve the final answer. 6 170175180185190 i-th Token 0 5 10 15 20 25 30 35 i-th Layer Before S-T 170175180185190 i-th Token After S-T 170175180185190 i-th Token Selected Span After − Before 10 −3 10 −3 −0.0004 −0.0002 0.0000 0.0002 0.0004 Figure 2 A concrete example of question-and-answer-to-context attention before and after S-T. Rows are layers and columns are context positions around the model-annotated span. The dashed vertical lines mark the selected training tokens. After S-T, attention to the selected span increases, while the change outside the span remains small. 4.2 Case study We next visualize how S-T changes the model’s use of the selected span. We compare question-and- answer-to-context attention before and after S-T, averaging over all heads and plotting the attention by layer. Figure 2 shows one such example. Before adaptation, the model already assigns some attention to the annotated evidence span, but the mass is sparse and uneven across layers. After training on that span, attention becomes stronger and more continuous around the selected tokens, especially in the middle layers. The difference panel shows that this change is localized: the warm region aligns with the training span, while most neighboring positions remain close to zero. This qualitative example suggests one mechanism behind S-T: adaptation on selected evidence induces a localized shift in attention toward tokens that are relevant to the current question. More visualized examples can be found in Appendix D. 4.3 Efficiency analysis T methods introduce extra overhead over direct inference because they add an adaptation stage before generation. We use pytorch FSDP (Paszke et al., 2019) for training and vLLM (Kwon et al., 2023) for inference. All measurements are conducted on a single NVIDIA H200 GPU. Figure 3 reports measured end-to-end latency normalized by full-context inference using Qwen3-4B-Thinking-2507. S-T has a higher latency in the beginning when the context length is relatively short. The crossover happens at longer context: S-T becomes cheaper than Full Context T from 64k onward on both benchmarks, and cheaper than Random Span T at 64k on LongBench-v2 and comparable on LongBench-Pro. Notably, at 128k context length, S-T has the lowest latency among the non-frozen-KV T methods. This is expected because Full Context T will incur significant overhead when the context length scales up as it trains on the entire input. Random Span T samples spans uniformly across the full context, which leads to a larger average effective training window of 0.50C, whereCis the context length. In contrast, model-annotated spans are more localized, resulting in shorter effective training windows on average (0.39Con LongBench-v2 and 0.37C on LongBench-Pro). The annotation cost dominates at shorter lengths, but quickly becomes smaller than the saved adaptation cost at long context. 5 Related Work Test-Time Training. T adapts model parameters to a single test input before prediction, using supervision derived from the input itself rather than from new labels (Sun et al., 2020). Earlier work studies when 7 16K32K64K128K Context Length 0× 2× 4× 6× 8× 10× 12× 14× Relative End-to-End Latency (Direct-Inference = 1x)↓ LongBench-v2 16K32K64K128K Context Length 1× 2× 3× 4× 5× 6× LongBench-Pro Base ModelLongLLMLinguaqTTT (frozen-KV)Random Span TTTFull Context TTTSelf-Guided T (ours) Figure 3 End-to-end latency normalized by full-context inference as context length increases using Qwen3-4B-Thinking- 2507. S-T incurs a higher latency in the beginning, but it becomes cheaper than other non-frozen KV cache T methods at longer context. self-supervised T helps or fails under distribution shift (Liu et al., 2021), and nearest-neighbor T adapts LLMs using retrieved examples at inference time (Hardt and Sun, 2024). More recent LLM work shows that per-instance adaptation can improve reasoning when the test input contains useful self-supervision (Akyürek et al., 2024). For long-context tasks, Bansal et al. (2026) show that T can be a more effective use of inference-time compute than simply generating more reasoning tokens. Related long-context T work also explores parameter-efficient adaptation for reasoning over long inputs (Chen et al., 2026a). These works primarily study how to perform adaptation efficiently. In this work, we study what tokens the model should be trained on at test time. S-T shows that selecting the right spans is a key component of effective long-context T. Long-Context LLMs. Modern LLMs increasingly support very long context windows, but a longer window does not guarantee reliable use of the information inside it. Models remain sensitive to evidence position, often degrading when relevant content appears in the middle of a long input (Liu et al., 2024), and long-context benchmarks such as LongBench, LongBench-v2, LongBench-Pro, ZeroSCROLLS, RULER, and HELMET make these failures visible across multi-document QA, code, dialogue, structured reasoning, recall, and long in-context learning tasks (Bai et al., 2024, 2025; Chen et al., 2026b; Shaham et al., 2023; Hsieh et al., 2024; Yen et al., 2025). A broad line of work addresses long-context limitations by extending usable context windows (Peng et al., 2024; Chen et al., 2024), improving prefill or attention efficiency (Jiang et al., 2024a), compressing prompts (Jiang et al., 2024b), retrieving external evidence (Lewis et al., 2020; Zhang et al., 2025b), or analyzing and steering attention behavior at inference time (Wu et al., 2025; Zhang et al., 2025b; Ye et al., 2026). These approaches largely aim to help the model condition on the right evidence or process long inputs more efficiently. We instead approach long-context reasoning from the perspective of test-time training: rather than compressing the long context or intervening in the decoding procedure, S-T only requires the model to first select relevant spans from the context before T. This keeps the model architecture and decoding algorithm unchanged, avoiding complex interventions that are often infeasible in modern inference engines, while yielding consistent gains. 6 Conclusion We propose Self-Guided T (S-T), a simple test-time adaptation framework for long-context LLMs that uses the model itself to select question-relevant evidence spans for training. Instead of adapting on the 8 full context or on randomly sampled spans, S-T first identifies supporting spans from the input context, adapts the model only on those selected spans, and then generates the final answer using the original full context. Our results on LongBench-v2 and LongBench-Pro show that S-T consistently improves over T on random span across Qwen3 and Llama-3.1 models, while remaining cheaper than other T variants at long context. Empirical results demonstrate that the effectiveness of long-context T depends critically on the quality of the test-time training tokens. Overall, S-T provides a simple yet effective framework for long-context test-time training, highlighting training-token selection as a promising direction for solving long-context tasks with T. We discuss future directions in Appendix E. References Ekin Akyürek, Mehul Damani, Adam Zweiger, Linlu Qiu, Han Guo, Jyothish Pari, Yoon Kim, and Jacob Andreas. The surprising effectiveness of test-time training for few-shot learning. arXiv preprint arXiv:2411.07279, 2024. Yushi Bai, Xin Lv, Jiajie Zhang, Hongchang Lyu, Jiankai Tang, Zhidian Huang, Zhengxiao Du, Xiao Liu, Aohan Zeng, Lei Hou, Yuxiao Dong, Jie Tang, and Juanzi Li. LongBench: A bilingual, multitask benchmark for long context understanding. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (ACL), 2024. https://arxiv.org/abs/2308.14508. Yushi Bai, Shangqing Tu, Jiajie Zhang, Hao Peng, Xiaozhi Wang, Xin Lv, Shulin Cao, Jiazheng Xu, Lei Hou, Yuxiao Dong, Jie Tang, and Juanzi Li. LongBench v2: Towards deeper understanding and reasoning on realistic long-context multitasks. In Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (ACL), 2025. https://aclanthology.org/2025.acl-long.183/. Rachit Bansal, Aston Zhang, Rishabh Tiwari, Lovish Madaan, Sai Surya Duvvuri, Fnu Devvrit, David Brandfonbrener, David Alvarez-Melis, Prajjwal Bhargava, Mihir Kale, and Samy Jelassi. Let’s (not) just put things in context: Test-time training for long-context LLMs. In The Fourteenth International Conference on Learning Representations, 2026. https://openreview.net/forum?id=H0bcEdPCoc. Yukang Chen, Shengju Qian, Haotian Tang, Xin Lai, Zhijian Liu, Song Han, and Jiaya Jia. LongLoRA: Efficient fine-tuning of long-context large language models. In International Conference on Learning Representations (ICLR), 2024. Zeming Chen, Angelika Romanou, Gail Weiss, and Antoine Bosselut. PERK: Long-context reasoning as parameter- efficient test-time learning. In The Fourteenth International Conference on Learning Representations, 2026a. https://openreview.net/forum?id=qxDTe8fIyA. Ziyang Chen, Xing Wu, Junlong Jia, Chaochen Gao, Qi Fu, Debing Zhang, and Songlin Hu. Longbench pro: A more realistic and comprehensive bilingual long-context evaluation benchmark. CoRR, abs/2601.02872, 2026b. Guhao Feng, Shengjie Luo, Kai Hua, Ge Zhang, Wenhao Huang, Di He, and Tianle Cai. In-place test-time training. In The Fourteenth International Conference on Learning Representations, 2026.https://openreview.net/forum?id= dTWfCLSoyl. Moritz Hardt and Yu Sun. Test-time training on nearest neighbors for large language models. In International Conference on Learning Representations (ICLR), 2024. https://openreview.net/forum?id=CNL2bku4ra. Cheng-Ping Hsieh, Simeng Sun, Samuel Kriman, Shantanu Acharya, Dima Rekesh, Fei Jia, Yang Zhang, and Boris Ginsburg. RULER: What’s the real context size of your long-context language models? In First Conference on Language Modeling, 2024. https://openreview.net/forum?id=kIoBbc76Sy. Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. LoRA: Low-rank adaptation of large language models. In International Conference on Learning Representations (ICLR), 2022. https://openreview.net/forum?id=nZeVKeeFYf9. Huiqiang Jiang, YUCHENG LI, Chengruidong Zhang, Qianhui Wu, Xufang Luo, Surin Ahn, Zhenhua Han, Amir H. Abdi, Dongsheng Li, Chin-Yew Lin, Yuqing Yang, and Lili Qiu. MInference 1.0: Accelerating pre-filling for long-context LLMs via dynamic sparse attention. In The Thirty-eighth Annual Conference on Neural Information Processing Systems, 2024a. https://openreview.net/forum?id=fPBACAbqSN. 9 Huiqiang Jiang, Qianhui Wu, Xufang Luo, Dongsheng Li, Chin-Yew Lin, Yuqing Yang, and Lili Qiu. LongLLMLingua: Accelerating and enhancing LLMs in long context scenarios via prompt compression. In Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (ACL), pages 1658–1677, 2024b.https: //aclanthology.org/2024.acl-long.91. Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph E. Gonzalez, Hao Zhang, and Ion Stoica. Efficient memory management for large language model serving with pagedattention. In Proceedings of the ACM SIGOPS 29th Symposium on Operating Systems Principles, 2023. Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, Sebastian Riedel, and Douwe Kiela. Retrieval-augmented generation for knowledge-intensive NLP tasks. In Advances in Neural Information Processing Systems (NeurIPS), volume 33, pages 9459–9474, 2020. https://arxiv.org/abs/2005.11401. Nelson F. Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang. Lost in the middle: How language models use long contexts. Transactions of the Association for Computational Linguistics, 12:157–173, 2024. doi: 10.1162/tacl_a_00638. https://aclanthology.org/2024.tacl-1.9/. Yuejiang Liu, Parth Kothari, Bastien van Delft, Baptiste Bellot-Gurlet, Taylor Mordan, and Alexandre Alahi. T++: When does self-supervised test-time training fail or thrive? In Advances in Neural Information Processing Systems (NeurIPS), pages 21808–21820, 2021.https://proceedings.neurips.c/paper/2021/hash/ b618c3210e934362ac261db280128c22-Abstract.html. Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. Pytorch: An imperative style, high-performance deep learning library. In Ad- vances in Neural Information Processing Systems 32, pages 8024–8035, 2019.http://papers.nips.c/paper/ 9015-pytorch-an-imperative-style-high-performance-deep-learning-library. Bowen Peng, Jeffrey Quesnelle, Honglu Fan, and Enrico Shippole. YaRN: Efficient context window extension of large language models. In The Twelfth International Conference on Learning Representations, 2024.https: //openreview.net/forum?id=wHBfxhZu1u. Qwen Team. Qwen3 technical report. arXiv preprint arXiv:2505.09388, 2025. https://arxiv.org/abs/2505.09388. Uri Shaham, Maor Ivgi, Avia Efrat, Jonathan Berant, and Omer Levy. ZeroSCROLLS: A zero-shot benchmark for long text understanding. In Findings of the Association for Computational Linguistics: EMNLP, pages 7977–7989, 2023. doi: 10.18653/v1/2023.findings-emnlp.536. https://aclanthology.org/2023.findings-emnlp.536. Yu Sun, Xiaolong Wang, Zhuang Liu, John Miller, Alexei A. Efros, and Moritz Hardt. Test-time training with self-supervision for generalization under distribution shifts. In Proceedings of the 37th International Conference on Machine Learning (ICML), pages 9229–9248, 2020. https://proceedings.mlr.press/v119/sun20b.html. Arnuv Tandon, Karan Dalal, Xinhao Li, Daniel Koceja, Marcel Rød, Sam Buchanan, Xiaolong Wang, Jure Leskovec, Sanmi Koyejo, Tatsunori Hashimoto, et al. End-to-end test-time training for long context. arXiv preprint arXiv:2512.23675, 2025. Llama Team. The llama 3 herd of models. CoRR, abs/2407.21783, 2024. Nandan Thakur, Nils Reimers, Andreas Rücklé, Abhishek Srivastava, and Iryna Gurevych. BEIR: A heterogeneous benchmark for zero-shot evaluation of information retrieval models. In Thirty-fifth Conference on Neural Infor- mation Processing Systems Datasets and Benchmarks Track (Round 2), 2021.https://openreview.net/forum?id= wCu6T5xFjeJ. Wenhao Wu, Yizhong Wang, Guangxuan Xiao, Hao Peng, and Yao Fu. Retrieval head mechanistically explains long-context factuality. In The Thirteenth International Conference on Learning Representations, 2025.https: //openreview.net/forum?id=EytBpUGB1Z. Xi Ye, Wuwei Zhang, Fangcong Yin, Howard Yen, and Danqi Chen. DySCO: Dynamic Attention-Scaling Decoding for Long-Context Language Models, April 2026. http://arxiv.org/abs/2602.22175. arXiv:2602.22175 [cs]. Howard Yen, Tianyu Gao, Minmin Hou, Ke Ding, Daniel Fleischer, Peter Izsak, Moshe Wasserblat, and Danqi 10 Chen. HELMET: How to evaluate long-context models effectively and thoroughly. In The Thirteenth International Conference on Learning Representations, 2025. https://openreview.net/forum?id=293V3bJbmE. Tianyuan Zhang, Sai Bi, Yicong Hong, Kai Zhang, Fujun Luan, Songlin Yang, Kalyan Sunkavalli, William T Freeman, and Hao Tan. Test-time training done right. arXiv preprint arXiv:2505.23884, 2025a. Wuwei Zhang, Fangcong Yin, Howard Yen, Danqi Chen, and Xi Ye. Query-focused retrieval heads improve long- context reasoning and re-ranking. In Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, and Violet Peng, editors, Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pages 23791–23805, Suzhou, China, November 2025b. Association for Computational Linguistics. ISBN 979-8-89176-332-6. doi: 10.18653/v1/2025.emnlp-main.1214. https://aclanthology.org/2025.emnlp-main.1214/. 11 Appendix A Implementation Details We use LoRA (Hu et al., 2022) for parameter-efficient test-time training. Following qTTT (Bansal et al., 2026), we apply LoRA only to the query projection layers, with rankr= 16 and scaling parameterα= 32. We optimize the LoRA parameters using AdamW with a 0.01 weight decay. For each method, we sweep the learning rate over3×10 −5 ,1×10 −4 ,3×10 −4 on a small validation set and select the best one for testing. For span annotation, LongBench-v2, which consists of multiple-choice questions, we append the answer choices to the question when prompting the model to annotate relevant spans. For LongBench-Pro, which contains open-ended questions, span annotation is performed using only the context and the question. B Annotation Coverage Table 4 Model annotation coverage on LongBench-v2 and LongBench-Pro. “Fallback” means the model fails to produce valid verbatim spans. ModelBenchmarkFallback Qwen3-4B-Thinking-2507 LongBench-v28.2% Qwen3-4B-Thinking-2507 LongBench-Pro 21.5% Llama-3.1-8B-InstructLongBench-v26.9% Llama-3.1-8B-InstructLongBench-Pro 39.9% Table 4 reports how often the model produces valid verbatim spans. Fallback instances use random spans, so they are equivalent to Random Span T for those cases. The fallback rate is low on LongBench-v2 for both base models, indicating that most instances receive genuine model-selected training spans. However, the fallback rate becomes higher on LongBench-Pro, especially for Llama-3.1-8B-Instruct, suggesting that self-annotation is more difficult on the open-ended benchmark. C Prompts Table 5 Qwen3-4B-Thinking-2507 prompt template for LongBench-v2. <|im_start|>system You are a helpful assistant. Read the context and answer the question.<|im_end|> <|im_start|>user CONTEXT QUESTION Pick from the following options: CHOICES Please show your choice in the answer field with only the choice letter, e.g., "answer": "C".<|im_end|> <|im_start|>assistant <think> 12 Table 6 Qwen3-4B-Thinking-2507 prompt template for LongBench-Pro. <|im_start|>system You are a helpful assistant. Read the context and answer the question.<|im_end|> <|im_start|>user CONTEXT QUESTION<|im_end|> <|im_start|>assistant <think> Table 7 Llama-3.1-8B-Instruct prompt template for LongBench-v2. <|begin_of_text|><|start_header_id|>system<|end_header_id|> Cutting Knowledge Date: December 2023 Today Date: 26 Jul 2024 <|eot_id|><|start_header_id|>user<|end_header_id|> Please read the following text and answer the question below. <text> CONTEXT </text> What is the correct answer to this question: QUESTION Choices: CHOICES Let’s think step by step. After thinking, choose a single, most likely answer. Output your final answer follows: "The correct answer is (insert choice here)".<|eot_id|><|start_header_id|>assistant<|end_header_id|> Table 8 Llama-3.1-8B-Instruct prompt template for LongBench-Pro. <|begin_of_text|><|start_header_id|>system<|end_header_id|> Cutting Knowledge Date: December 2023 Today Date: 26 Jul 2024 You are a helpful assistant. Read the context and answer the question.<|eot_id|><|start_header_id|>user<|end_header_id|> CONTEXT QUESTION<|eot_id|><|start_header_id|>assistant<|end_header_id|> 13 D Qualitative Examples 190019201940 i-th Token 0 5 10 15 20 25 30 35 i-th Layer Before S-T 190019201940 i-th Token After S-T 190019201940 i-th Token Selected Span After − Before 10 −4 10 −3 10 −4 10 −3 −0.0004 −0.0002 0.0000 0.0002 0.0004 Figure 4 Example of span attention before and after S-T. 190019502000 i-th Token 0 5 10 15 20 25 30 35 i-th Layer Before S-T 190019502000 i-th Token After S-T 190019502000 i-th Token Selected Span After − Before 10 −4 10 −4 −0.00010 −0.00005 0.00000 0.00005 0.00010 Figure 5 Example of span attention before and after S-T. E Future Directions T opens up opportunities for adapting LLMs in realistic production settings. In many applications, users may upload a long document, such as a financial report, legal contract, or book, and then ask multiple questions about it. Unlike methods that require architectural changes or specialized attention mechanisms, T relies on standard gradient-based adaptation and on-the-fly weight updates. This makes it compatible with modern training and serving infrastructure, where each conversation session could maintain its own lightweight adapted weights for multi-turn use, personalization, or document-specific specialization. However, the largest bottleneck is latency: even parameter-efficient T adds adaptation overhead before generation, which needs to be reduced. Addressing these challenges is an important direction for making T a practical framework for solving long-context tasks in real-world systems. 14 1895190019051910 i-th Token 0 5 10 15 20 25 30 35 i-th Layer Before S-T 1895190019051910 i-th Token After S-T 1895190019051910 i-th Token Selected Span After − Before 10 −4 10 −3 10 −4 10 −3 −0.0003 −0.0002 −0.0001 0.0000 0.0001 0.0002 0.0003 Figure 6 Example of span attention before and after S-T. 15