Carlos's Debrief

June 19, 2026 04:00
44ArXiv Papers
28Web Findings
72Total Sources
Topics searched: Artificial Intelligence Machine Learning Large Language Models Security & Cybersecurity Cryptography Zero Knowledge Quantum Computing Crypto & Blockchain AI Agents & Reasoning AI Safety & Alignment
June 19, 2026 04:00 — Curated by Hermes Research Scout

⚡ Quick Summary

📰 News Headlines

🗡️ Relevant to your Katana work 2

📄 ArXiv Papers

🧠 Agentic / Sovereign / Execution 3

Alaia Solko-Breslin, Pramod Kaushik Mudrakarta, Mihai Christodorescu, Somesh Jha, Krishnamurthy Dj Dvijotham
Securing AI agents that operate in complex digital environments has become a critical need, and runtime monitoring approaches that formulate and enforce policies expressed in a formal language like…
cs.CRcs.AI
Reza Soosahabi, Vivek Namsani
Agentic AI systems increasingly rely on language-model components to interpret instructions, process external data, invoke tools, and coordinate with other agents. These capabilities make…
cs.CRcs.AI
Jun He, Deying Yu
Autonomous agents are increasingly connected to cloud, deployment, and data-control workflows, but production mutation authority should not reside inside non-deterministic reasoning processes.…
cs.CRcs.AIcs.DC

🧠 How / Instructions / Shape 2

Nityanand Mathur, Hamees Sayed, Wasim Madha, Apoorv Singh, Sameer Khurana, Akshat Mandloi, Sudarshan Kamath
Style-captioned text-to-speech systems use natural language to control voice characteristics, but how individual words influence acoustic output remains unclear. Understanding this is critical for…
cs.AI
Harshit Singh, Ayush Pratap Singh, Nityanand Mathur
Flow-matching text-to-speech systems achieve remarkable zero-shot quality but remain static after deployment: pronunciation errors on out-of-vocabulary proper nouns persist unless the model is…
cs.AI

🧠 Multi-lcb / Extending / Livecodebench 2

Bercan Turkmen, Vyas Raina
Malware analysts often inspect compiled binaries through decompiled pseudo-C, when source code is unavailable. Recent work suggests that large language models (LLMs) can assist this process by…
cs.CRcs.AI
Maria Ivanova, Pavel Zadorozhny, Rodion Levichev, Ivan Petrov, Adamenko Pavel, Ivan Lopatin, Alexey Kutalev, Dmitrii Babaev
LiveCodeBench (LCB) has recently become a widely adopted benchmark for evaluating large language models (LLMs) on code-generation tasks. By curating competitive programming problems, constantly…
cs.AIcs.PL

🧠 Egocentric / Video / Uniego 2

Wenhao Chi, Arkaprava Sinha, Dominick Reilly, Hieu Le, Srijan Das
Egocentric video understanding is inherently limited by the narrow perspective of wearable cameras: a single viewpoint, a single modality, a single model cannot capture the full richness of human…
cs.CVcs.LG
Juncheng Ma, Jianxin Bi, Yufan Deng, Xuanran Zhai, Kewei Zhang, Ye Huang, Bo Liang, Shukai Gong, Jiankai Tu, Xiaotian Tang, Jiaxin Li, Kaiqi Chen, Duomin Wang, Yuqi Wang, Bingyi Kang, Eric Huang, Zhiyang Dou, Zhen Dong, Enze Xie, Wojciech Matusik, Tat-Seng Chua, Daquan Zhou
Embodied foundation models are expected to benefit from data scaling like large language models, but face a much tighter data bottleneck. Teleoperated real-robot trajectories remain the dominant…
cs.CV

🧠 Privacy / Predictability / Fine-grained 2

Shanghao Shi, Chaoyu Zhang, Heng Jin, Yang Xiao, Yevgeniy Vorobeychik, William Yeoh, Ning Zhang, Y. Thomas Hou, Wenjing Lou
Federated learning (FL) enables multiple parties to collaboratively fine-tune language models for domain-specific tasks without sharing raw data. Since full model fine-tuning is often prohibitively…
cs.CR
Linda Lu, Karthik Sridharan
Differential privacy (DP) ensures rigorous individual-level privacy guarantees against even the most knowledgeable attackers, but its worst-case nature can impose a costly privacy-accuracy tradeoff.…
cs.LG

🧠 How / Transparent / Diffusiongemma 1

Joshua Engels, Callum McDougall, Bilal Chughtai, Janos Kramar, Senthoran Rajamanoharan, Cindy Wu, Arthur Conmy, Asic Q Chen, Jean Tarbouriech, Min Ma, Brendan O'Donoghue, João Gabriel Lopes de Oliveira, Rohin Shah, Neel Nanda
LLM reasoning transparency is a critical affordance for understanding model decisions, mitigating misuse and misalignment, and debugging surprising model behaviors. However, DiffusionGemma performs a…
cs.LGcs.AI

🧠 Structuring / Tokenizing / Distributed 1

Ruizhong Qiu, Yinglong Xia, Dongqi Fu, Hanqing Zeng, Ren Chen, Xiangjun Fan, Hong Li, Hong Yan, Hanghang Tong
Generative recommendation is an emerging paradigm that has shown promise in industrial recommendation systems, aiming to predict users' next interactions from their historical behaviors. At the core…
cs.IRcs.AI

🧠 Toward / Calibrated / Mixture-of-experts 1

Gina Wong, Drew Prinster, Suchi Saria, Rama Chellappa, Anqi Liu
Calibration aligns a model's predictive uncertainty with the frequencies of its empirical outcomes and is important for understanding and trusting reported probabilities. Recent work shows that…
cs.AIcs.LG

🧠 Ledgeragent / Structured / State 1

Md Nayem Uddin, Amir Saeidi, Eduardo Blanco, Chitta Baral
Policy-adherent tool-calling agents in customer-service domains must maintain task states across turns while calling tools and obeying domain policies. Task states consist of relevant facts,…
cs.AIcs.CL

🧠 Deepswip / Quotient-wmc / Counterfactuals 1

Saimun Habib, Vaishak Belle, Fengxiang He
Neurosymbolic systems such as DeepProbLog combine neural perception with probabilistic logic, but standard inference is associational. Counterfactual reasoning additionally requires a causal…
cs.AI

🧠 Sarlo-80 / Worldwide / Slant 1

Solène Debuysère, Nicolas Trouvé, Nathan Letheule, Elise Colin, Georgia Channing
Multimodal foundation models have advanced rapidly thanks to large optical benchmarks, but comparable resources for synthetic aperture radar (SAR) remain limited. Existing SAR--optical datasets…
cs.CVcs.AIcs.DB

🧠 Optimal / Deterministic / Multicalibration 1

Georgy Noarov, Aaron Roth
A model is multicalibrated on a collection of group weights $G$ if it is calibrated -- i.e. unbiased even conditional on its prediction -- not just overall, but also after reweighting contexts by…
cs.LGmath.STstat.ML

🧠 Token / Group / Element 1

Przemyslaw Musialski
We place the attention token on the group: a token is an element $g_i$ of a matrix Lie group $G$ -- a bare transformation, with no feature payload and no external action $ρ(g)$ carrying it. To our…
cs.LGcs.CVcs.GR

🧠 Multi-task / Bayesian / In-context 1

Qingyang Zhu, Eric Karl Oermann, Kyunghyun Cho
Bayesian predictive inference provides a principled framework for uncertainty quantification, data efficiency, and robust generalization. However, exact inference is often intractable, and scalable…
cs.LG

🧠 Execution-state / Capsules / Graph-bound 1

Liang Su
Mainstream LLM serving systems reuse prefix work mainly through paged or radix key-value (KV) caches. This is highly effective for high-throughput, high-concurrency serving, but it manages only one…
cs.LGcs.DC

🧠 Probe-and-refine / Tuning / Repository 1

Asa Shepard, Jeannie Albrecht
LLM-based coding agents need higher-level operational knowledge about a repository (which files house which subsystems, how to run the test suite, which workflows have historically led to wrong…
cs.SEcs.LG

🧠 Memorywam / Efficient / World 1

Sizhe Yang, Juncheng Mu, Tianming Wei, Chenhao Lu, Xiaofan Li, Linning Xu, Zhengrong Xue, Zhecheng Yuan, Dahua Lin, Jiangmiao Pang, Huazhe Xu
Robust robotic manipulation in the real world requires not only an understanding of the current observation, but also memory and dynamics modeling. World action models (WAMs) possess these…
cs.RO

🧠 Timeprove / Propose / Then 1

Arkaprava Sinha, Dominick Reilly, Siddharth Krishnan, Hieu Le, Srijan Das
Long Video Question Answering (LVQA) requires identifying sparse, query-relevant evidence within hours-long untrimmed videos. Existing approaches either process videos densely with large…
cs.CV

🧠 Observation / Electroweak / Production 1

CMS Collaboration
The first evidence of electroweak (EW) production of pairs of Z bosons in association with two jets (jj) in the final state ZZjj $\to$ $\ell\ellνν$jj, where $\ell$ = e, $μ$, is reported by the CMS…
hep-ex

🧠 Thinking / Boxes / Editing 1

Pradhaan S Bhat, Naveen Chandra R, Rishubh Parihar, Vaibhav Vavilala, R. Venkatesh Babu, D. A. Forsyth, Anand Bhattad
Text and 2D-conditioning interfaces provide weak, ambiguous control over spatial transformations in image editing -- particularly under large object motions and camera changes. Prior work has used 3D…
cs.CV

🧠 Incorporating / Physical / Source 1

Mateusz J. Mróz, Radosław Poleski, Andrzej Udalski, Jan Skowron, Paweł Pietrukowicz, Michał K. Szymański, Przemek Mróz, Mariusz Gromadzki, Patryk Iwanek, Szymon Kozłowski, Milena Ratajczak, Krzysztof A. Rybicki, Dorota M. Skowron, Igor Soszyński, Krzysztof Ulaczyk, Marcin Wrona, Zofia Buzik
Modeling of complex microlensing events suffers from many difficult-to-disentangle degeneracies. This is especially the case for orbital motion of the source in a binary system, the so-called…
astro-ph.SRastro-ph.IM

🧠 Calibration / Without / Comprehension 1

Arastoo Zibaeirad, Marco Vieira
Whether LLMs scoring well on vulnerability benchmarks genuinely reason about security or merely pattern-match on contaminated data remains unresolved. We present CWE-Trace, a framework for LLM…
cs.CRcs.AIcs.SE

🧠 A-compass / Formal / Foundations 1

Tamara Tagliavia, Silvia Ghilezan
In the information age, one of the leading problems is how to ensure individual's privacy. Depending on the context in which privacy is considered, various data privacy models have emerged. However,…
cs.CRcs.LO

🧠 Image / Encryption / Algorithm 1

Ans Ibrahim, Fadhil Abbas Fadhil, Mahameed Reza Feizi Derakhshi, Maryam Mahdi Alhusseini, Nikolai Safiullin
The paper proposes a dynamic approach to image encryption, combining the use of Convolutional Neural Networks (CNNs) and classical cryptography to improve the security and flexibility of image…
cs.CRcs.SE

🧠 Invariants / Colored / Braid 1

Illia E. Rohozhkin
In this paper, a braid is regarded as a dynamical system of points in the plane. The states of this dynamical system are given by Delaunay triangulations. This construction makes it possible to…
math.GN

🧠 Bioeth-beacon / Confidential / On-chain 1

Christos Galanopoulos, Kimon Antonios Provatas, Ilias Georgakopoulos-Soares
The Global Alliance for Genomics and Health (GA4GH) Beacon protocol lets researchers ask whether a genomic variant has been observed in a participating cohort and receive aggregate variant-level…
q-bio.GNcs.CR

🧠 Low-cost / Multi-precision / Systolic 1

George Alexakis, Dimitrios Schoinianakis, Giorgos Dimitrakopoulos
Fully Homomorphic Encryption (FHE) ensures robust data privacy but suffers from prohibitive computational overhead. Accelerating FHE on AI hardware like Tensor Processing Units (TPUs) is promising,…
cs.CR

🧠 Disarm / Target / Electronic 1

Tasneem Suha, Tanzim Mahfuz, Rima Asmar Awad, Prabuddha Chakraborty
Program runtime or timing attacks exploit variations in a program's execution times to extract sensitive information from the program (e.g. encryption keys, sensitive variable data, intellectual…
cs.CR

🧠 Renormalization / Group / Flow 1

Kevin T. Grosvenor, Subodh P. Patil
In this paper, we study the statistical field-theoretic renormalization of active flocks via the MSRDJ action formulation for stochastic systems, focusing on the Toner-Tu theory of `Malthusian…
cond-mat.softhep-th

🧠 Caching / Dollars / Not 1

Madhulatha Mandarapu, Sandeep Kunkunuru
When a cache miss fetches from cloud object storage, the bill is per GET request and per byte of egress, not latency. Classic caching minimizes the miss rate, the wrong objective: a rarely but…
cs.DBcs.DS

🧠 Benchmark / Quantum / Algorithms 1

Daniel Molpeceres, Sirui Lu, J. Ignacio Cirac, Barbara Kraus
We compare the performance of representative cooling, adiabatic, and optimization algorithms for ground-state preparation in the presence of noise. Using an exactly solvable family of quadratic…
quant-ph

🧠 Generating / Robot / Hands 1

Sha Yi, Nicklas Hansen, Xueqian Bai, Carmelo Sferrazza, Michael T. Tolley, Xiaolong Wang
Robot learning has advanced rapidly in learning control, but learning the physical body of a robot remains much more difficult because jointly searching over design and control creates a very large…
cs.RO

🧠 Topological / Codes / Space 1

Chong-Yuan Xu, Ze-Chuan Liu, Yong Xu
Topological codes form one of the most important classes of stabilizer codes. Most existing algebraic constructions and analyses of topological codes assume translation invariance. Here we show that…
quant-phcond-mat.quant-gas

🧠 Caltennis / Multi-view / Tennis 1

Ilona Demler, Xinran Xie, Blake Werner, Anna Szczuka, Pietro Perona
The Caltech Tennis Dataset (CalTennis) is a large-scale video benchmark for evaluating monocular-to-3D pose estimation in the wild. CalTennis comprises over 11 million frames (51 hours) of tennis…
cs.CV

🧠 Fid / Lottery / Quantifying 1

Nicolas Dufour, Alexei A. Efros, Patrick Pérez
The Frechet Inception Distance (FID) is the de facto arbiter of image generation, yet most papers report just a single number from a single trained model using a single sampling seed. How…
cs.CV

🧠 Current / World / Lack 1

Jinpeng Lu, Dexu Zhu, Haoyuan Shi, Linghan Cai, Guo Tang, Yinda Chen, Jie Cao, Duyu Tang, Yi Zhang, Yong Dai, Xiaozhu Ju
World models are increasingly regarded as a decisive step toward artificial general intelligence, yet modeling the physical world demands more than rendering convincing frames on demand: it requires…
cs.CV

🧠 Janusmesh / Fast / Zero-shot 1

Siang-Ling Zhang, Huai-Hsun Cheng, Tsung-Ju Yang, Yu-Lun Liu
Creating 3D visual illusions, a single 3D mesh that reveals entirely different semantics from various viewing angles, is a fascinating but tough challenge. Existing optimization-based methods are…
cs.CV

🧠 Ssd / Spatially / Speculative 1

Shilong Xiang, Zirui Zhang, Lijun Yu, Chengzhi Mao
Autoregressive models excel in visual generation by treating images as 1D sequences of discrete tokens, mirroring language modeling. However, this flattening discards the intrinsic 2D spatial…
cs.CV

🌐 Web Findings

🤗 HuggingFace Papers 20

Achieving dexterous robotic manipulation in the real world heavily relies on human supervision and algorithm engineering, which becomes a central bottleneck in the pursuit of…
▲ 6HuggingFace Papers
Conditional diffusion and flow models routinely fail to satisfy the very constraints that define their task. For instance, a depth-conditioned model often produces images whose…
▲ 4HuggingFace Papers
Visual thinking should not only sound right; it should show its evidence. While recent vision-language models (VLMs) can produce natural-language reasoning traces, these traces…
▲ 4HuggingFace Papers
Agent benchmarks are growing fast, but no single benchmark touches more than four or five of the dimensions that deployment exposes. This paper aggregates the largest coordinated…
▲ 17HuggingFace Papers
Real-world spatial intelligence requires reasoning over a continuous and evolving 3D world, yet existing VLMs and tool-augmented agents largely remain tied to static, stateless…
▲ 20HuggingFace Papers
Test-time reasoning is increasingly used as a serving-time control knob, but extra reasoning is not uniformly valuable: it can repair failed attempts, waste compute on…
HuggingFace Papers
Current AI-driven game development has made substantial progress in asset generation, gameplay design, and web-based game coding, yet project-level code engineering on…
HuggingFace Papers
World Action Models (WAMs) commonly rely on video generation to bridge visual world modeling and robot control. However, video-based WAMs face three coupled limitations: dense…
▲ 1HuggingFace Papers
Multi-step LLM pipelines fail through interactions among retrieval, reasoning, and formatting steps, so prompt-only optimization can miss bottlenecks in the chain. We present FAPO…
▲ 2HuggingFace Papers
Typical video object-centric learning (VOCL) approaches employ slot-based frameworks that rely on reconstruction-driven encoder-decoder architectures, where learning is mediated…
▲ 2HuggingFace Papers
Current agentic robot systems can write executable Code-as-Policy programs, observe feedback, and revise behavior across multiple attempts, but they remain largely task-driven:…
▲ 27HuggingFace Papers
Hybrid linear attention models offer an appealing path to faster long-context inference: they reduce the quadratic cost and KV-cache burden of full softmax attention while…
▲ 1HuggingFace Papers
Style-content dual-reference generation aims to synthesize an image that preserves the structure and semantics of a content reference while adopting the style of a separate style…
▲ 14HuggingFace Papers
Accurate mechanical properties (or materials) Young's modulus (E), Poisson's ratio (ν) and density (ρ) are essential for reliable physics simulation of digital worlds, but most 3D…
▲ 2HuggingFace Papers
Video world models are moving toward preserving an observed world under controllable camera and object motion while allowing its environmental state to change. Yet these controls…
▲ 1HuggingFace Papers
Recent retrieval-augmented generation (RAG) approaches have demonstrated strong capability in handling complex queries, yet current research overlooks a critical challenge:…
▲ 4HuggingFace Papers
Precise 3D spatial orchestration in text-to-video generation remains a significant challenge, particularly for multi-object scenes where semantic layout and temporal dynamics are…
HuggingFace Papers
While 10B-level industrial foundation models have pushed the boundaries of image inpainting, their prohibitive computational costs severely hinder practical deployment.…
▲ 34HuggingFace Papers
Advances in radiance fields have enabled photorealistic novel view synthesis. In several domains, large-scale real-world datasets have been developed to support comprehensive…
▲ 4HuggingFace Papers
Dexterous interaction with articulated objects is important for household, assistive, and humanoid manipulation, where multi-finger hands can provide compliant contact patterns…
▲ 8HuggingFace Papers

📝 OpenReview 2

The advancements in Large Language Models (LLMs) have been hindered by their substantial sizes, which necessitate LLM compression methods for practical deployment. Singular Value…
OpenReview
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-scale image data and…
OpenReview

🦞 Lobste.rs 1

A startup out of Europe built an AI system that matches Mythos on zero-day discovery, using widely available models, even air-gapped. You've probably never heard of it. Here's the…
Lobste.rs

🎓 Google Scholar 5

This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantitative analysis of…
Google Scholar

🔗 All Sources

  1. [1] How Transparent is DiffusionGemma?
  2. [2] Structuring and Tokenizing Distributed User Interest Context for Generative Recommendation
  3. [3] Toward Calibrated Mixture-of-Experts Under Distribution Shift
  4. [4] How Do Instructions Shape Speech? Cross-Attention Attribution for Style-Captioned…
  5. [5] LedgerAgent: Structured State for Policy-Adherent Tool-Calling Agents
  6. [6] DeepSWIP: Quotient-WMC Counterfactuals for Neural Probabilistic Logic Programs
  7. [7] SARLO-80: Worldwide Slant SAR Language Optic Dataset 80cm
  8. [8] Sovereign Execution Brokers: Enforcing Certificate-Bound Authority in Agentic Control…
  9. [9] FlowEdit: Associative Memory for Lifelong Pronunciation Adaptation in Flow-Matching TTS
  10. [10] Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages
  11. [11] UNIEGO: Proxies as Mediators for Unified Egocentric Video Representation Learning
  12. [12] Optimal Deterministic Multicalibration and Omniprediction
  13. [13] The Token Is a Group Element: On Lie-Algebra Attention over Matrix Lie Groups
  14. [14] Predictability as a Fine-Grained Measure for Privacy
  15. [15] Multi-Task Bayesian In-Context Learning
  16. [16] Execution-State Capsules: Graph-Bound Execution-State Checkpoint and Restore for…
  17. [17] Probe-and-Refine Tuning of Repository Guidance for Coding Agents
  18. [18] MemoryWAM: Efficient World Action Modeling with Persistent Memory
  19. [19] TimeProVe: Propose, then Verify for Efficient Long Video Temporal Reasoning in Activities…
  20. [20] Observation of electroweak production of pairs of Z bosons in proton-proton collisions at…
  21. [21] Thinking in Boxes: 3D Editing in Real Images Made Easy
  22. [22] Incorporating physical source parameters into microlensing modeling
  23. [23] From Efficiency to Leakage -- Privacy Backdoor in Federated Language Model Fine-Tuning
  24. [24] Efficient and Sound Probabilistic Verification for AI Agents
  25. [25] Calibration Without Comprehension: Diagnosing the Limits of Fine-Tuning LLMs for…
  26. [26] A-COMPASS: Formal Foundations for Anonymity Analysis in Microdata
  27. [27] Analyzing Defensive Misdirection Against Model-Guided Automated Attacks on Agentic AI…
  28. [28] Image Encryption Algorithm Based on Convolutional Neural Networks and Dynamic S-Box…
  29. [29] Multi-View Decompilation for LLM-Based Malware Classification
  30. [30] Invariants of the Colored Braid Groupoid
  31. [31] bioETH-Beacon: A Confidential On-Chain Genomic Beacon with Encrypted Counts, Filters, and…
  32. [32] Low-Cost Multi-Precision Systolic Arrays for Accelerating FHE NTTs on AI ASICs
  33. [33] DISARM: Target Electronic Device Informed Mitigation of Software Runtime Side-Channel…
  34. [34] On the Renormalization Group Flow of Active Flocks
  35. [35] Caching for Dollars, Not Hits: An Exact Offline Reference for Cloud-Egress Caching and…
  36. [36] Benchmark of quantum algorithms for ground state preparation in the presence of noise
  1. [37] Generating Robot Hands from Human Demonstrations
  2. [38] Topological Codes Based on Space Groups
  3. [39] CalTennis: Large Multi-View Tennis Video Dataset and Benchmark of Monocular-to-3D Pose…
  4. [40] The FID Lottery: Quantifying Hidden Randomness in Generative-Model Evaluation
  5. [41] HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining
  6. [42] Current World Models Lack a Persistent State Core
  7. [43] JanusMesh: Fast and Zero-Shot 3D Visual Illusion Generation via Cross-Space Denoising
  8. [44] SSD: Spatially Speculative Decoding Accelerates Autoregressive Image Generation
  9. [45] "Mythos" at Home, and It's Called AISLE
  10. [46] Taylor-Calibrate: Principled Initialization for Hybrid Linear Attention Distillation
  11. [47] FlowBender: Feedback-Aware Training for Self-Correcting Conditional Flows
  12. [48] DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated…
  13. [49] DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis
  14. [50] Playful Agentic Robot Learning
  15. [51] S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence
  16. [52] Understanding the Behaviors of Environment-aware Information Retrieval
  17. [53] FAPO: Fully Autonomous Prompt Optimization of Multi-Step LLM Pipelines
  18. [54] Selective Synergistic Learning for Video Object-Centric Learning
  19. [55] Adaptive Volumetric Mechanical Property Fields Invariant to Resolution
  20. [56] ENPIRE: Agentic Robot Policy Self-Improvement in the Real World
  21. [57] Holo-World: Unified Camera, Object and Weather Control for Video World Model
  22. [58] FreeStyle: Free Control of Style-Content Dual-Reference Generation from Community LoRA…
  23. [59] JAMER: Project-Level Code Framework Dataset and Benchmark on Professional Game Engines
  24. [60] ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?
  25. [61] Beyond Static Leaderboards: Predictive Validity for the Evaluation of LLM Agents
  26. [62] Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance
  27. [63] Thinking with Visual Grounding
  28. [64] LooseControlVideo: Directorial Video Control using Spatial Blocking
  29. [65] Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning
  30. [66] Catalyst breakthroughs in methane dry reforming: Employing machine learning for future…
  31. [67] Advancements in Artificial Intelligence: Breakthroughs, Challenges and the Road Ahead
  32. [68] Large language models (LLM) in computational social science: prospects, current state,…
  33. [69] Artificial intelligence and machine learning in cybersecurity: a deep dive into…
  34. [70] Quantum machine learning: A comprehensive review of integrating AI with quantum computing…
  35. [71] SVD-LLM: Truncation-aware Singular Value Decomposition for Large Language Model…
  36. [72] Diffusion$^2$: Dynamic 3D Content Generation via Score Composition of Video and…