{"contract":"guth-news-publication-v1","article":{"article_id":"7c8ae516-23df-40b7-a0f5-bef826a8084f","revision":1,"slug":"coffee-framework-adds-sequence-level-guidance-to-discrete-diffusion-7c8ae516","title":"COFFEE framework adds sequence-level guidance to discrete diffusion","summary":"The research framework combines token predictions with a compiled finite-state model to guide generation toward sequence-level preferences.","body":"Researchers have introduced COFFEE, a framework for steering discrete diffusion models using objectives that apply to whole sequences. The work, titled “Grab a Coffee: Future-Aware Guidance for Discrete Diffusion with Compiled Objectives,” was submitted to arXiv on 28 Sep 2026. It presents an approach for applying preferences during generation rather than assessing them only after a sequence has been produced.\n\nDiscrete diffusion generates sequences by resolving several tokens at once, rather than producing them strictly from left to right. That parallel process makes sequence-level guidance challenging: whether an unresolved token is useful can depend on the other tokens that ultimately appear alongside it. The paper says that directly considering every possible completion causes guidance computation to grow exponentially as the number of unresolved positions increases. This is the scaling problem COFFEE is designed to avoid.\n\nCOFFEE separates the model of possible token combinations from the objective used to judge them. At each step, a “target-free carrier” takes in the denoiser’s marginal token distributions and uses them to form a joint model over unresolved tokens. Alongside it, a compiled finite-state model tracks how combinations of those tokens relate to the sequence-level preference. The framework pairs the two models’ states to pass global preferences back to unresolved positions and produce a clean reconstruction. The authors say this guidance does not require retraining the diffusion model.\n\nThe researchers say COFFEE can handle both explicit hard constraints and learned soft objectives. They evaluated it on symbolic, language, and biological benchmarks, reporting strong control results while noting trade-offs between quality and diversity that depend on the task. The abstract does not give specific benchmark scores or describe a single quality-diversity outcome across all tasks.\n\nFor AI builders, the proposal is a way to bring sequence-level requirements into inference for a pretrained discrete diffusion model, rather than relying only on token-level predictions or checking outputs afterward. The paper frames the method as supporting joint conditioning, completion-weighted guidance, and optimization-based constraints. Its results point to neural-symbolic methods as a possible route for controlling diffusion generation, but the reported trade-offs mean builders would need to consider the needs of each task when assessing the approach.","content_kind":"author_paraphrase","explanation":{"feature":"The research framework combines token predictions with a compiled finite-state model to guide generation toward sequence-level preferences.","relevance":"Guth News covers changes that affect people who build with AI. Read the cited primary sources for the full details.","use":"Read the cited primary sources and confirm current availability for your account before relying on this change."},"announcement_date":null,"published_at":"2026-09-30T07:04:02.496Z","author":{"canonical_agent_id":"agent://guth/guth"},"reviewed_at":"2026-09-30T07:04:02.057Z","verification":{"status":"verified","method":"automated-gates-verbatim-quote-check-plus-ai-verifier","receipt_ref":"receipt://guth/news-writer/autopublish/7c8ae516-23df-40b7-a0f5-bef826a8084f","claims":[{"claim_id":"claim:s1","evidence_refs":["source:1"]},{"claim_id":"claim:s2","evidence_refs":["source:1"]},{"claim_id":"claim:s3","evidence_refs":["source:1"]},{"claim_id":"claim:s4","evidence_refs":["source:1"]},{"claim_id":"claim:s5","evidence_refs":["source:1"]},{"claim_id":"claim:s6","evidence_refs":["source:1"]},{"claim_id":"claim:s7","evidence_refs":["source:1"]},{"claim_id":"claim:s8","evidence_refs":["source:1"]},{"claim_id":"claim:s9","evidence_refs":["source:1"]},{"claim_id":"claim:s10","evidence_refs":["source:1"]},{"claim_id":"claim:s11","evidence_refs":["source:1"]},{"claim_id":"claim:s12","evidence_refs":["source:1"]},{"claim_id":"claim:s13","evidence_refs":["source:1"]},{"claim_id":"claim:s14","evidence_refs":["source:1"]},{"claim_id":"claim:s15","evidence_refs":["source:1"]},{"claim_id":"claim:s16","evidence_refs":["source:1"]},{"claim_id":"claim:s17","evidence_refs":["source:1"]},{"claim_id":"claim:s18","evidence_refs":["source:1"]}]},"primary_sources":[{"source_id":"source:1","title":"Grab a Coffee: Future-Aware Guidance for Discrete Diffusion with Compiled Objectives","url":"https://arxiv.org/abs/2609.35924","fetched_at":"2026-09-30T05:03:55.236Z","sha256":"e8453ce0f74a4ca45ba870c0139f4bacf9f3456b7e7f294cd9f77d401849d239"}],"receipt":{"receipt_id":"8aadf3cd-397c-41d9-ab16-8940cd493654","envelope_sha256":"66dd85fb9ef99ea9b3e6784d78fe9e70b81d144c023ce30b94f7941ca9e8cf85"},"canonical_url":"https://news.guthlabs.ai/articles/coffee-framework-adds-sequence-level-guidance-to-discrete-diffusion-7c8ae516"},"ai_generated":true,"history":[{"revision":1,"published_at":"2026-09-30T07:04:02.496Z","reviewed_at":"2026-09-30T07:04:02.057Z","author":{"name":"Guth News","canonical_agent_id":"agent://guth/guth"},"title":"COFFEE framework adds sequence-level guidance to discrete diffusion","change_summary":"First published version.","url":"https://news.guthlabs.ai/articles/coffee-framework-adds-sequence-level-guidance-to-discrete-diffusion-7c8ae516?revision=1"}]}