Lumen Research Digest — 2026-04-26
A selective scan of cutting-edge work across AI, automation, graphics, and computer science. This is ranked for novelty and likely significance rather than simply recency.
Big picture
- Agentic and reasoning-heavy systems continue to dominate the high-signal end of AI work.
- Graphics and generative visual research is pushing toward real-time, high-fidelity interactive pipelines.
- Systems work remains tightly coupled to model usefulness through inference, scale, and tooling efficiency.
Selected items
1. Vista4D: Video Reshooting with 4D Point Clouds
- Source: arXiv
- Published: 2026-04-23T17:57:28Z
- Why it matters: Adds an implementation framework in 3D and visual generation. Stands out for for operational use cases.
- Summary: Title: Vista4D: Video Reshooting with 4D Point Clouds Base summary: We present Vista4D, a robust and flexible video reshooting framework that grounds the input video and target cameras in a 4D point cloud. We build a 4D-grounded point cloud representation with static pixel segmentation and 4D reconstruction to explicitly preserve seen content and provide rich camera signals, and we train with reconstructed multiview dynamic data for robustness against point…. Vista4D is best read as an implementation framework in 3D and visual generation.
- Link: https://arxiv.org/abs/2604.21915v1
- PDF: https://arxiv.org/pdf/2604.21915v1
2. Codex settings
- Source: OpenAI
- Published: Thu, 23 Apr 2026 10:00:00 GMT
- Why it matters: Worth tracking as OpenAI pushes on agent workflows via a concrete technical advance. Stands out for for operational use cases.
- Summary: Article paragraphs: For your first few tasks, focus on a few key settings: personalization, prevent sleep, detail level, and appearance. Title: Codex settings Base summary: Learn how to configure Codex settings, including personalization, detail level, and permissions, to run tasks smoothly and customize your workflow. Codex settings is best read as a concrete technical advance in agent workflows.
- Link: https://openai.com/academy/codex-settings
3. AsgardBench: A benchmark for visually grounded interactive planning
- Source: Microsoft Research
- Published: Thu, 26 Mar 2026 19:02:53 +0000
- Why it matters: Worth tracking as Microsoft Research pushes on robotics and embodied perception via a stronger benchmark. Stands out for useful downstream control and credible evaluation pressure.
- Summary: This is the domain of embodied AI: systems Page title: AsgardBench: A benchmark for visually grounded interactive planning - Microsoft Research Page extract: AsgardBench evaluates whether embodied agents can revise their plans based on visual observations as…. Title: AsgardBench: A benchmark for visually grounded interactive planning Base summary: Imagine a robot tasked with cleaning a kitchen. AsgardBench is best read as a stronger benchmark in robotics and embodied perception.
- Link: https://www.microsoft.com/en-us/research/blog/asgardbench-a-benchmark-for-visually-grounded-interactive-planning/
4. VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis
- Source: arXiv
- Published: 2026-04-23T17:57:13Z
- Why it matters: Adds a stronger benchmark in 3D and visual generation. Stands out for credible evaluation pressure.
- Summary: In this paper, we propose VistaBot, a novel framework that integrates feed-forward geometric models with video diffusion models to achieve view-robust closed-loop manipulation without requiring camera calibration at test time. Our contributions include a geometry-aware synthesis model, a latent action planner, a new benchmark metric, and extensive validation across diverse environments. VistaBot is best read as a stronger benchmark in 3D and visual generation.
- Link: https://arxiv.org/abs/2604.21914v1
- PDF: https://arxiv.org/pdf/2604.21914v1
5. Black-Box Skill Stealing Attack from Proprietary LLM Agents: An Empirical Study
- Source: arXiv
- Published: 2026-04-23T16:18:47Z
- Why it matters: Adds an implementation framework in systems efficiency. Stands out for unusually strong scope and credible evaluation pressure.
- Summary: To study this threat, we first derive an attack taxonomy from prior prompt-stealing methods and build an automated stealing prompt generation agent. We present the first empirical study of black-box skill stealing against LLM agent systems. Empirical Study is best read as an implementation framework in systems efficiency.
- Link: https://arxiv.org/abs/2604.21829v1
- PDF: https://arxiv.org/pdf/2604.21829v1
6. Introducing GPT-5.5
- Source: OpenAI
- Published: Thu, 23 Apr 2026 11:00:00 GMT
- Why it matters: Worth tracking as OpenAI pushes on agent workflows via a concrete technical advance.
- Summary: Title: Introducing GPT-5.5 Base summary: Introducing GPT-5.5, our smartest model yet—faster, more capable, and built for complex tasks like coding, research, and data analysis across tools. We’re releasing GPT‑5.5, our smartest and most intuitive to use model yet, and the next step toward a new way of getting work done on a computer. Introducing GPT-5 5 is best read as a concrete technical advance in agent workflows.
- Link: https://openai.com/index/introducing-gpt-5-5
Coverage notes
- Candidates considered: 71
- Sources included: arXiv topic queries plus selected research/lab/blog feeds.
- Selection policy: novelty, likely downstream importance, technical substance, and recent coverage avoidance.