Работы о diffusion LM за 2026 год

Не полный каталог: здесь показаны профильные публикации, сохранённые в публичном корпусе. Число работ не измеряет качество или важность метода.

UTC: .

Профильных работ по текущим фильтрам: 4.

По месяцам и темам
МесяцРабот
Январь0
Февраль0
Март0
Апрель0
Май0
Июнь0
Июль0
Август0
Сентябрь4

Темы могут пересекаться: одна работа учитывается в каждой своей теме, но только один раз в общем числе. Связь с темой не доказывает качество или воспроизводимость.

Свежие работы · Методика отбора

2026-09-07T07:55:40+00:00 · Дообучение · Источник

Оригинальное название: In-Place Instruction Following in Diffusion Language Models

Оригинальная аннотация: Diffusion Large Language Models (dLLMs) generate text via bidirectional iterative denoising, naturally supporting user-specified constraints anchored at arbitrary output positions, a paradigm known as In-place Prompting (IPP). We formalize this as the In-place Instruction Following (IIF) task and construct IIF-Bench, a hierarchical benchmark spanning literal, style, and discourse-function constraints, paired with a rubric-based local-global evaluation protocol. An inference-time attention-bias probe suggests that vanilla dLLMs often under-prioritize constraint spans during denoising. We then p

Полезно для: Подход применим к диффузионным большим языковым моделям, поддерживающим пользовательские ограничения в произвольных позициях вывода. · Ограничение: При денойзинге обычные dLLM часто недостаточно приоритизируют участки с ограничениями.

2026-09-06T01:09:14+00:00 · без темы · Источник

Оригинальное название: A Ticket from Marginals to Joints: Coupled-Noise Distillation for One-Step Block Generation in Diffusion Language Models

Оригинальная аннотация: Autoregressive language models commit one token per forward pass; diffusion language models commit a block of tokens over several steps. We ask whether a block can be committed in a single forward pass. We study this with a noise-conditioned masked denoiser: a data-independent Gaussian noise field is added to the mask embeddings so that, in principle, each sampled field selects one joint mode of the block. The established way of training such a model is to sample several fields per example and let them compete for the data, by winner-take-all or importance weighting. This gives the noise only

Полезно для: Метод применим к генерации блоков за один шаг с одним прямым проходом на блок. · Ограничение: Источник не приводит количественных результатов и не уточняет размеры протестированных моделей.

2026-09-02T04:52:01+00:00 · без темы · Источник

Оригинальное название: Predict, Don't Iterate: Efficient Adaptive-Length Infilling for Diffusion Language Models

Оригинальная аннотация: Diffusion language models (DLMs) have emerged as a promising alternative to the auto-regressive paradigm. With bidirectional attention and any-order generation, DLMs naturally fit infilling tasks, which require generating a middle span conditioned on both the prefix and the suffix. However, infilling is sensitive to the length of the span, while DLMs require the length to be fixed before generation. Although prior studies extend DLMs to dynamic lengths, they still suffer from two limitations. (i) Sensitivity to initial length. These methods require a preset length to initialize the search and

Полезно для: Задачи инфиллинга для диффузионных языковых моделей, включая бенчмарки кода и текста. · Ограничение: В источнике отсутствуют подробности реализации алгоритма и результаты по отдельным бенчмаркам.

2026-09-01T08:07:21+00:00 · Дообучение · Источник

Оригинальное название: Membership Inference in Fine-tuned Diffusion Language Models via Token-level Memorization Asymmetry

Оригинальная аннотация: Diffusion language models (DLMs) have recently emerged as an alternative modeling paradigm to autoregressive LMs, offering advantages such as parallel generation and bidirectional context modeling. Despite growing interest in their generative capabilities, the privacy risks of DLMs remain underexplored. We identify a phenomenon termed token-level memorization asymmetry through theoretical analysis of diffusion training dynamics. Building on this finding, we propose Q-Skew, a quantile-weighted skewness-based indicator for membership inference on finetuned DLMs. Experiments across multiple fine-

Полезно для: Не указано · Ограничение: Не указано

Даты публичного снимка
Сбор, зафиксированный в снимке
Самая новая публикация по данным снимка
Построение снимка

Это сохранённые сведения, а не время последней попытки сборщика или гарантия полноты корпуса.