Transporting Task Vectors across Different Architectures without Training
Researchers introduced Theseus, a training-free method to transport task-specific parameter updates across large language models with different architectures.
Search signals, briefings, benchmarks and glossary terms.
Researchers introduced Theseus, a training-free method to transport task-specific parameter updates across large language models with different architectures.
Preprints and research are separated from the executive feed. Publication here is not validation; open the paper and inspect its evidence.
| Date ↓ | Source | Development | Posture | Horizon | Topics |
|---|---|---|---|---|---|
| 31 Jul 2026 | arXiv cs.LG — Machine Learning | Transporting Task Vectors across Different Architectures without Training ↗ | Investigate | Next 12 months | model adaptation, fine tuning, model optimization |
| 31 Jul 2026 | arXiv cs.LG — Machine Learning | STEREODISCO: Discovering Stereotypicality in LLMs ↗ | Investigate | Next 12 months | model bias, responsible ai, model evaluation |
| 31 Jul 2026 | arXiv cs.LG — Machine Learning | Doubly Robust Functional Representation Learning for Longitudinal Causal Inference with Irregular Histories ↗ | Investigate | Next 12 months | causal inference, longitudinal studies, irregular time series |
| 31 Jul 2026 | arXiv cs.LG — Machine Learning | Uncertainty quantification for trustworthy deep learning: Methods and measures ↗ | Investigate | Next 12 months | uncertainty quantification, model risk, explainability |
| 31 Jul 2026 | arXiv cs.LG — Machine Learning | Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale ↗ | Investigate | Next 12 months | agentic ai, synthetic data, model training |
| 31 Jul 2026 | arXiv cs.LG — Machine Learning | Beyond the Bidirectional Promise: Re-evaluating the Robustness of Diffusion Language Models ↗ | Investigate | Next 12 months | model evaluation, llm security, model robustness |
| 29 Jul 2026 | arXiv cs.CL — Computation and Language | Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models ↗ | Investigate | Next 12 months | llm security, model evaluation, safety alignment |
| 29 Jul 2026 | arXiv cs.CL — Computation and Language | VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation ↗ | Investigate | Next 12 months | multimodal reasoning, rag, visual llm |
| 29 Jul 2026 | arXiv cs.CL — Computation and Language | Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore ↗ | Investigate | Next 12 months | rag, model evaluation, explainability |
| 29 Jul 2026 | arXiv cs.CL — Computation and Language | KletterMix: Climbing Toward High-Quality German Pretraining Data - The Full Report ↗ | Investigate | Next 12 months | pretraining data, german language models, corpus development |
| 29 Jul 2026 | arXiv cs.CL — Computation and Language | Ranked by Position: Order Sensitivity as an Exploitable Attack Surface in LLM Listwise Recommenders ↗ | Investigate | Next 12 months | llm security, model risk, recommendation systems |
| 29 Jul 2026 | arXiv cs.CL — Computation and Language | Measuring and Improving Behavioral Consistency in Large Language Models through Fact-Heuristic-Emotion State Enforcement ↗ | Investigate | Next 12 months | model evaluation, model risk, explainability |
The default view excludes research papers, removes low-confidence items, sorts by publication date and caps each publisher at four displayed items. Research has its own view, capped at six papers per research feed. Every headline opens the underlying source.
One development is evidence, not momentum. The board does not label a topic “rising” from a single article, and the narrative implications are explicitly marked as interpretive assessments. Use repeated, independent sources over time before treating a topic as a trend.
Receive eight source-linked developments at 06:30 UK, with factual summary kept separate from interpretive assessment.
Free. Daily at 06:30 UK. Unsubscribe with one click.