Explorer
- .. (Parent Directory)
- pdfs/
- a-lightweight-hybrid-transformer-crf-architecture-for-multi-type-bangl-2605.25463.md
- a-predictive-law-for-on-policy-self-distillation-from-world-feedback-2605.30070.md
- adapting-multilingual-embedding-models-to-turkish-via-cross-lingual-to-2605.29992.md
- adaptive-teacher-exposure-for-self-distillation-in-llm-reasoning-2605.11458.md
- agentic-cost-aware-query-planning-with-knowledge-distillation-for-big-2605.17831.md
- amr-sd-asymmetric-meta-reflective-self-distillation-for-token-level-cr-2605.18529.md
- apertus-llm-family-expansion-via-distillation-and-quantization-2605.29128.md
- are-diffusion-language-models-good-database-analysts-2605.27791.md
- are-full-rollouts-necessary-for-on-policy-distillation-2605.31490.md
- avsd-adaptive-view-self-distillation-by-balancing-consensus-and-teache-2605.20643.md
- backtracking-when-it-strays-mitigating-dual-exposure-biases-in-llm-rea-2605.19433.md
- balancing-knowledge-distillation-for-imbalance-learning-with-bilevel-o-2605.17839.md
- bridging-reasoning-trajectories-in-on-policy-distillation-via-near-fut-2606.00305.md
- cepo-rlvr-self-distillation-using-contrastive-evidence-policy-optimiza-2605.19436.md
- clinseekagent-automating-multimodal-evidence-seeking-for-agentic-clini-2605.20176.md
- clore-content-level-optimization-for-reasoning-efficiency-2605.22211.md
- codeskill-learning-self-evolving-skills-for-coding-agents-2605.25430.md
- consisguard-aligning-safety-deliberation-with-policy-enforcement-in-ll-2605.31073.md
- context-instrumental-data-distillation-for-kubernetes-manifest-generat-2605.25835.md
- cornerstones-or-stumbling-blocks-deciphering-the-rock-tokens-in-on-pol-2605.09253.md
- cross-paradigm-knowledge-distillation-a-comprehensive-study-of-bidirec-2605.19299.md
- decoupling-kl-and-trajectories-a-unified-perspective-for-sft-dagger-of-2605.16826.md
- defermem-query-time-evidence-distillation-via-reinforcement-learning-f-2605.22411.md
- diffusion-large-language-models-for-visual-speech-recognition-2605.28456.md
- diffusionblocks-2506.14202.md
- diladiff-distilled-latent-augmented-diffusion-for-language-modeling-2605.23605.md
- distilling-llm-feedback-for-lean-theorem-proving-2605.30861.md
- distilling-tabular-foundation-models-for-structured-health-data-2605.18702.md
- divergence-decoding-inference-time-unlearning-via-auxiliary-models-2605.31293.md
- edge-opd-internalizing-privileged-context-with-evidence-guided-on-poli-2605.23493.md
- extreme-region-policy-distillation-2605.25582.md
- fast-tensorization-of-neural-networks-via-slice-wise-feature-distillat-2605.19842.md
- fedsdr-federated-self-distillation-with-rectification-2605.18028.md
- forgetting-has-neighbors-localized-collateral-forgetting-in-machine-un-2605.31317.md
- foster-first-order-dataset-distillation-for-text-based-sequential-reco-2605.30772.md
- from-fact-overwriting-to-knowledge-evolution-causal-editing-via-on-pol-2605.28303.md
- gdsd-reinforcement-learning-as-guided-denoiser-self-distillation-for-d-2605.29398.md
- hint-sd-targeted-hindsight-self-distillation-for-long-horizon-agents-2605.17873.md
- it-takes-two-complementary-self-distillation-for-contextual-integrity-2605.20258.md
- knowledge-graph-driven-expert-level-reasoning-for-neuroscience-2605.25183.md
- llamion-technical-report-2605.25676.md
- looped-diffusion-language-models-2605.26106.md
- mimo-multilingual-information-retrieval-via-monolingual-objectives-2605.31171.md
- mira-mid-training-rubric-anchoring-for-source-aware-data-selection-2605.30288.md
- mta-multi-granular-trajectory-alignment-for-large-language-model-disti-2605.01374.md
- multi-objective-learning-for-diffusion-models-a-statistical-theory-und-2605.25210.md
- omniopd-logit-free-on-policy-distillation-via-speculative-verification-2606.01476.md
- opd-rethinking-the-advantage-design-for-on-policy-distillation-2606.01039.md
- pair-in-pair-out-latent-multi-token-prediction-for-efficient-llms-2605.27255.md
- pocket-foundation-models-distilling-tfms-into-cpu-ready-gradient-boost-2605.18654.md
- post-trained-moe-can-skip-half-experts-via-self-distillation-2605.18643.md
- post-training-is-about-states-not-tokens-a-state-distribution-view-of-2605.22731.md
- prune-opd-efficient-and-reliable-on-policy-distillation-for-long-horiz-2605.07804.md
- restoring-the-sweet-spot-pass-rate-weighted-self-distillation-for-llm-2605.27765.md
- rosd-reflective-on-policy-self-distillation-for-language-model-reasoni-2605.28014.md
- sae-fd-sparse-autoencoder-feature-distillation-for-continual-learning-2605.25525.md
- safeguarding-text-to-image-generative-models-against-unauthorized-know-2605.22060.md
- same-evidence-different-answers-canonical-context-on-policy-distillati-2605.30251.md
- sd-search-on-policy-hindsight-self-distillation-for-search-augmented-r-2605.18299.md
- search-e1-self-distillation-drives-self-evolution-in-search-augmented-2605.22511.md
- self-policy-distillation-via-capability-selective-subspace-projection-2605.22675.md
- self-supervised-on-policy-distillation-for-reasoning-language-models-2605.17497.md
- sgmd-score-gradient-matching-distillation-for-few-step-video-diffusion-2605.30116.md
- shield-a-diverse-clinical-note-dataset-and-distilled-small-language-mo-2605.03301.md
- shred-retain-set-free-unlearning-via-self-distillation-with-logit-demo-2605.07482.md
- simct-recovering-lost-supervision-for-cross-tokenizer-on-policy-distil-2605.07711.md
- skill-conditioned-gated-self-distillation-for-llm-reasoning-2605.28791.md
- slimqwen-exploring-the-pruning-and-distillation-in-large-moe-model-pre-2605.08738.md
- sometin-beta-pass-notin-sbpn-improving-multilingual-asr-for-nigerian-l-2605.17710.md
- sra-span-representation-alignment-for-large-language-model-distillatio-2605.01205.md
- strong-teacher-not-needed-on-distillation-in-llm-pretraining-2605.23857.md
- subliminal-learning-is-steering-vector-distillation-2606.00995.md
- tailoring-teaching-to-aptitude-direction-adaptive-self-distillation-fo-2605.22263.md
- teacher-guided-policy-optimization-for-on-policy-reasoning-distillatio-2605.13230.md
- textteacher-what-can-language-teach-about-images-2605.22098.md
- the-bridge-garden-dilemma-in-llm-distillation-why-mixing-hard-and-soft-2605.26246.md
- the-distillation-game-adaptive-attacks-efficient-defenses-2605.22737.md
- thinkswitch-context-distillation-with-lora-and-weight-interpolation-fo-2606.01080.md
- towards-distillation-guarantees-under-algorithmic-alignment-for-combin-2605.20074.md
- towards-generalization-of-block-attention-via-automatic-segmentation-a-2605.15913.md
- tracer-persistent-regularization-for-robust-multimodal-finetuning-2605.29380.md
- triplet-block-diffusion-rwkv-2605.25969.md
- trust-region-behavior-blending-for-on-policy-distillation-2605.31159.md
- trust-region-on-policy-distillation-2606.01249.md
- uni-opd-unifying-on-policy-distillation-with-a-dual-perspective-recipe-2605.03677.md
- unveil-unified-visual-textual-integration-and-distillation-for-multi-m-2605.24530.md
- variance-reduction-for-expectations-with-diffusion-teachers-2605.21489.md
- vcap-hypergeometric-rewards-for-weak-to-strong-visual-captioning-2605.28023.md
- vision-opd-learning-to-see-fine-details-for-multimodal-llms-via-on-pol-2605.18740.md
- what-and-when-to-distill-selective-hindsight-distillation-for-multi-tu-2605.19447.md
- what-makes-a-strong-model-a-unified-spectral-analysis-of-knowledge-tra-2606.01292.md
- when-are-teacher-tokens-reliable-position-weighted-on-policy-self-dist-2605.21606.md
- when-confidence-misleads-suffix-anchoring-and-anchor-proximity-confide-2605.28181.md
- x-token-projection-guided-cross-tokenizer-knowledge-distillation-2605.21699.md
- xmodel-kd-cross-modal-knowledge-distillation-for-3d-scene-perception-u-2605.30111.md