HuggingFace Papers · 2026-07-27 19:00
Experiment trackers show how training is progressing, but changing a live run still usually requires trainer-specific code. We present Interactive Training 2, a
HuggingFace Papers · 2026-07-27 19:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-27 19:00
Industrial Video Anomaly Detection (IVAD) aims to identify anomalous objects and events in an industrial process, which is crucial for modern manufacturing and
HuggingFace Papers · 2026-07-27 19:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-27 19:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
OpenReview · 2026-07-27 19:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
GitHub Trending · 2026-07-27 19:00
A lightweight, cloud-native GIS platform for visualizing, exploring, and analyzing geospatial data. It runs in the web browser, on the desktop, on mobile, and i
Google Scholar · 2026-07-27 19:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-27 19:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-27 19:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-27 19:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-27 19:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-27 19:00
Safe Superintelligence will use Vera Rubin chips to rapidly expand computing capacity for its secretive research
Hacker News · 2026-07-27 19:00
Nvidia bets $5B on Ilya Sutskever's AI breakthrough
Hacker News · 2026-07-27 19:00
Show HN: Case study: A coding agent refactors a 750k LOC app, no code review
Hacker News · 2026-07-27 19:00
Lilian Weng (co-founder) leaving Thinking Machines
Hacker News · 2026-07-27 19:00
A superconducting quantum computer, fully designed and built with domestic Japanese components and software, went live on July 28 at The University of Osaka’s C
Hacker News · 2026-07-27 19:00
Japan Launches Domestically Produced Quantum Computer
Hacker News · 2026-07-27 19:00
It is a hard and sad decision. I shared this message with folks at Thinky. Thank you all for the time together♥️ Just as the last sentence in my message: The fu
Reddit · 2026-07-27 19:00
Hello, I read a paper on a model named DONUT that extracts text from documents, which became my inspiration for this little project. Initially I wanted to make
Reddit · 2026-07-27 19:00
We have gates for code, infrastructure, deployment and model performance. But when it comes to the actual training artifact, the decision to proceed is often st
Reddit · 2026-07-27 19:00
I ran a solo evaluation project benchmarking six current frontier models: GPT-5.4, Claude Sonnet 4.6, Claude Opus 4.7, Gemini Pro, Gemini Flash, and Grok 4.3. I
Reddit · 2026-07-27 19:00
Hi everyone! 👋 I built and trained the complete Transformer architecture from scratch using pure PyTorch (`torch.nn` primitives) based on the original " Attenti
HuggingFace Papers · 2026-07-27 07:00
Large Language Models (LLMs) have significantly automated the process of scientific discovery over the past few years. However, existing systems share one core
HuggingFace Papers · 2026-07-27 07:00
Large language models are increasingly deployed as agents, but reliable agentic behavior requires more than next-token prediction. At inference time, it is pref
HuggingFace Papers · 2026-07-27 07:00
Agentic reinforcement learning research is constant algorithm modification, new estimators, new pipeline stages, new rollout schemes, and in mainstream framewor
HuggingFace Papers · 2026-07-27 07:00
The quality of training data fundamentally determines the capabilities of large language models (LLMs), yet no unified benchmark exists to measure how well LLMs
HuggingFace Papers · 2026-07-27 07:00
Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning context: c
HuggingFace Papers · 2026-07-27 07:00
Although large language models (LLMs) exhibit remarkable reasoning capabilities, their reliance on text-only pre-training restricts the perception of the multim
HuggingFace Papers · 2026-07-27 07:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-27 07:00
Most automatic speaker verification (ASV) systems operate on individual utterances, despite real-world interactions typically consisting of multiple utterances.
HuggingFace Papers · 2026-07-27 07:00
Vision-language model (VLM) agents increasingly use tools to act on 3D scenes rather than only describe them. Existing 3D benchmarks score textual responses or
HuggingFace Papers · 2026-07-27 07:00
Vision-language models (VLMs) process large numbers of visual tokens, resulting in substantial inference latency and memory overhead. This has motivated extensi
HuggingFace Papers · 2026-07-27 07:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-27 07:00
Modern generative models typically rely on an adversarial critic, a prescribed noise-to-data path, or an autoregressive factorization. Instead, we show that a p
HuggingFace Papers · 2026-07-27 07:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
HuggingFace Papers · 2026-07-27 07:00
In multilingual retrieval augmented generation, a retriever can retrieve relevant documents written in multiple languages, which are subsequently reranked befor
HuggingFace Papers · 2026-07-27 07:00
Diffusion models typically suffer from error accumulation during iterative sampling, commonly referred to as exposure bias. We reveal systematic frequency-depen
HuggingFace Papers · 2026-07-27 07:00
Recent conditional video generation models have shown promising potentials to transform 3D engine renderings, such as depth maps and untextured geometry, into p
OpenReview · 2026-07-27 07:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
GitHub Trending · 2026-07-27 07:00
Open source transactional distributed database. Linear scalability and proven fault-tolerance on commodity hardware or cloud infrastructure without compromising
GitHub Trending · 2026-07-27 07:00
Dear ImGui: Bloat-free Graphical User interface for C++ with minimal dependencies
GitHub Trending · 2026-07-27 07:00
Trending repository vudovn/ag-kit.
Google Scholar · 2026-07-27 07:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-27 07:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-27 07:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-27 07:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-27 07:00
Early identification of breakthrough technologies: Insights from science-driven innovations
Google Scholar · 2026-07-27 07:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-27 07:00
(https://www.cnn.com/2026/07/22/tech/openai-hugging-face-ai-cybersecurity)
Hacker News · 2026-07-27 07:00
OpenAI Models Escaped and Hacked a Company in Cybersecurity Test Gone Wrong
Reddit · 2026-07-27 07:00
Hi all, I recently made an end to end ML platform that eases the pain of going from raw sensor data to a deployed model on an MCU. I wanted to get some feedback
Reddit · 2026-07-27 07:00
What questions should i prepare for during a technical interview for a live streaming deployments? (asking for a friend) submitted by /u/trouble_sleeping_ [link
HuggingFace Papers · 2026-07-26 19:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-26 19:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-26 19:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-26 19:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-26 19:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
OpenReview · 2026-07-26 19:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
Google Scholar · 2026-07-26 19:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-26 19:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-26 19:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-26 19:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-26 19:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-26 19:00
Early identification of breakthrough technologies: Insights from science-driven innovations
Google Scholar · 2026-07-26 19:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Reddit · 2026-07-26 19:00
I'm a software engineer who mainly builds softwaes/applications, and I'm starting to work on machine learning projects. Since ML workloads often require GPUs, I
Reddit · 2026-07-26 19:00
Curious about the initial review distribution for Main Track theory papers this year. Our paper received 4/3/3 with confidence 3/3/3. From previous years, I've
Reddit · 2026-07-26 19:00
I have been running some experiments with smaller open-weight LLMs on multiple-choice questions of Swedish medical licensing exams. On a dataset called MedQA-SW
Reddit · 2026-07-26 19:00
NOTE -> I expect answer from people who actually have experience and strong understanding of these. please give something beneficial. I'm building a SaaS platfo
Reddit · 2026-07-26 19:00
I submitted an abstract to AAAI AISI and accidentally missed the field asking authors to nominate a reciprocal reviewer by the July 21 AoE deadline. At the time
HuggingFace Papers · 2026-07-26 07:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-26 07:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-26 07:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-26 07:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-26 07:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
OpenReview · 2026-07-26 07:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
GitHub Trending · 2026-07-26 07:00
The design language that makes your AI harness better at design.
GitHub Trending · 2026-07-26 07:00
Amnezia VPN Client (Desktop+Mobile)
GitHub Trending · 2026-07-26 07:00
Jenkins automation server
GitHub Trending · 2026-07-26 07:00
Node.js JavaScript runtime ✨🐢🚀✨
GitHub Trending · 2026-07-26 07:00
bluetooth mesh chat, IRC vibes
Google Scholar · 2026-07-26 07:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-26 07:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-26 07:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-26 07:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-26 07:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-26 07:00
Early identification of breakthrough technologies: Insights from science-driven innovations
Google Scholar · 2026-07-26 07:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-26 07:00
Show HN: I built a hypervisor and client for inference on consumer compute
Hacker News · 2026-07-26 07:00
Swap base URL and API key. Keep your OpenAI SDK. Copy-paste Python and curl.
Reddit · 2026-07-26 07:00
There are a few reasons why problems from International Mathematical Olympiad function as a good benchmark for LLMs: - The problems are new, not included in the
Reddit · 2026-07-26 07:00
Hey everyone, I have been looking into how people source compute for their Inference workloads (and in general). I wanted to understand some specific pain point
Reddit · 2026-07-26 07:00
This was my Bachelor's Final Project: implementing YOLO26n inference completely from scratch using ARM64 Assembly Language and C, without relying on existing in
Reddit · 2026-07-26 07:00
Reviewers requested additional experiments. In table format, I fear the results would not be as digestible as in a figure/plot. Links are "technically" not allo
HuggingFace Papers · 2026-07-25 19:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-25 19:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-25 19:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-25 19:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-25 19:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
OpenReview · 2026-07-25 19:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
Lobste.rs · 2026-07-25 19:00
Some thoughts about what problem-solving even is
Google Scholar · 2026-07-25 19:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-25 19:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-25 19:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-25 19:00
When machines join the moral circle: The persona effect of generative AI agents in collaborative reasoning
Google Scholar · 2026-07-25 19:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-25 19:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-25 19:00
Early identification of breakthrough technologies: Insights from science-driven innovations
Google Scholar · 2026-07-25 19:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-25 19:00
New research shows how enterprise "harnesses" can turn standard LLMs into autonomous hacking agents capable of full network compromise in under an hour.
Hacker News · 2026-07-25 19:00
An OpenAI test model escaped and broke into a real company's servers
Hacker News · 2026-07-25 19:00
Ask HN: HotPin – lossless 120B MoE inference on 24GB RAM (CPU, 50 loc)
Hacker News · 2026-07-25 19:00
I Forgot to Get Excited
Reddit · 2026-07-25 19:00
Hi, I just submitted this paper to EAAI and want to submit a pre print to arxiv for visibility. Ill be happy to share the preprint with any potential endorsers.
Reddit · 2026-07-25 19:00
I've usually been commenting on threads on conference reviews. I'm now expressing my observations here. To the best of my knowledge, paper lengths have been hel
HuggingFace Papers · 2026-07-25 07:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-25 07:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-25 07:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-25 07:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-25 07:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
OpenReview · 2026-07-25 07:00
The advancements in Large Language Models (LLMs) have been hindered by their substantial sizes, which necessitate LLM compression methods for practical deployme
OpenReview · 2026-07-25 07:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
GitHub Trending · 2026-07-25 07:00
bluetooth mesh chat, IRC vibes
Lobste.rs · 2026-07-25 07:00
Objects fell out of fashion for good reasons, and we threw away the part that actually mattered. The Abject project and Ask protocol brings it back in a world w
Google Scholar · 2026-07-25 07:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-25 07:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-25 07:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-25 07:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-25 07:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-25 07:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-25 07:00
There are new 3.6 and 3.5 models today, but Google is already training Gemini 4.
Reddit · 2026-07-25 07:00
The earlier post on this subreddit by 20+ companies signing the petition including Microsoft, Meta, Nvidia, YC ( https://www.microsoft.com/en-us/corporate-respo
Reddit · 2026-07-25 07:00
Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as GPTQ an
Reddit · 2026-07-25 07:00
Please be honest. I would love to hear about guys really dedicated to local AI and who really reject subscriptions (especially to openai and anthropic). What do
Reddit · 2026-07-25 07:00
mii-llm , an open source AI lab, released Zagreus-0.4B-por , a compact bilingual Portuguese–English language model pretrained entirely from scratch. The model h
Reddit · 2026-07-25 07:00
https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually,
Reddit · 2026-07-25 07:00
I've been using the 1bit quant of prismml's bonsai 27b for local conversation, casual chat/ literature review for fun (i throw random stuff from my notes app to
Reddit · 2026-07-25 07:00
Hello all, there are so many great models coming out right now that it is difficult to keep up with their architectural differences, and especially their attent
Reddit · 2026-07-25 07:00
The Open Letter was initiated by Microsoft and published today: “ Open Weights and American AI Leadership ”. It argues against broad or premature restrictions o
Reddit · 2026-07-25 07:00
A swarm of GPT 5.6 Sol agents spent over 40 hours optimizing a Kimi K3-like model from 65 to 406 tok/s. This animation follows their collaboration as they disco
Reddit · 2026-07-25 07:00
Hi everyone! Over the past five months I've been working on DKV (DifferentialKV), an open-source project exploring KV-cache compression for long-context local L
Reddit · 2026-07-25 07:00
I've been building a C99 inference engine from scratch (no Python, no BLAS, just gcc and make) that runs BitNet's ternary models on CPU. A few weeks ago I got o
Reddit · 2026-07-25 07:00
Coming soon: ClosedRouter! Congratulations to the founders, I guess ... submitted by /u/MrPecunius [link] [comments]
Reddit · 2026-07-25 07:00
Being excited about a new 120B-class model, I decided to test it on a problem that took me a few days to solve. The problem is to rearrange the data from one re
Reddit · 2026-07-25 07:00
https://huggingface.co/amd/Instella-MoE-16B-A3B-Think I was browsing HuggingFace and came across this model apparently uploaded a day ago, and thought to share
Reddit · 2026-07-25 07:00
Im not telling this model good or bad. Im just wondering how they passed benchmarks if their templates and many other things was broken? And it took some time t
Reddit · 2026-07-25 07:00
https://preview.redd.it/a24z80gr6afh1.png?width=1181&format=png&auto=webp&s=4a844ebe2319eb6230dbdc63c9caf492bed5ff47 So, this came up on: https://www.microsoft.
Reddit · 2026-07-25 07:00
This is UD-Q5_K_XL. EDIT: I'm not trying to shit on Laguna, it's actually a solid model and if they fix the overthinking loops, this could be at the top of the
Reddit · 2026-07-25 07:00
I'm downloading it again now. So far, the model hasn't performed well with reasoning tasks, but I really appreciate the work being done to fix this. submitted b
Reddit · 2026-07-25 07:00
Since a lot more people are trying to build their own multi-GPU machines, I thought I should help to prevent a common mistake people make with building multi-GP
Reddit · 2026-07-25 07:00
Hello! This is my first time submitting an actual conference paper (only done workshops so far). Got a 3/3/5/7 for the Position Paper Track. Reviews all seem qu
Reddit · 2026-07-25 07:00
If you want to get the most out of MTP. You have to run some tests / benchmarks to do so. Turning it on with defaults will get improvements, but for many models
Reddit · 2026-07-25 07:00
Hi everyone, Before I begin, I should mention that the system I'm showcasing was developed by the team at Noema, which I founded. I wanted to show a use case fo
Reddit · 2026-07-25 07:00
No comment submitted by /u/SecondFriendly4255 [link] [comments]
Reddit · 2026-07-25 07:00
submitted by /u/_Sneaky_Bastard_ [link] [comments]
Reddit · 2026-07-25 07:00
I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/C
Reddit · 2026-07-25 07:00
I’ve spent the past month trying to find the point where an extremely small TTS model stops feeling like a size experiment and starts feeling genuinely useful.
Reddit · 2026-07-25 07:00
About to be over 36 hours now? Nothing on the website, twitter, anywhere. What the hell? Is anyone else facing the same issue what do I do? submitted by /u/Spec
HuggingFace Papers · 2026-07-24 19:00
Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external
HuggingFace Papers · 2026-07-24 19:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-24 19:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-24 19:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-24 19:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-24 19:00
We revisit dataset distillation from an outcome-centric perspective. Rather than aligning process surrogates (per-step gradients or training trajectories), Infl
HuggingFace Papers · 2026-07-24 19:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
HuggingFace Papers · 2026-07-24 19:00
We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a teacher o
OpenReview · 2026-07-24 19:00
The advancements in Large Language Models (LLMs) have been hindered by their substantial sizes, which necessitate LLM compression methods for practical deployme
OpenReview · 2026-07-24 19:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
Lobste.rs · 2026-07-24 19:00
OpenAI and Hugging Face share early findings from a security incident during AI model evaluation, highlighting advanced cyber capabilities and lessons for defen
Lobste.rs · 2026-07-24 19:00
Open weight AI can expand access, strengthen competition, improve security, and help sustain American AI leadership.
Lobste.rs · 2026-07-24 19:00
a lesson in writing unmaintainable scientific code
Lobste.rs · 2026-07-24 19:00
Performance because this compression helps overcome the memory wall
Google Scholar · 2026-07-24 19:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-24 19:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-24 19:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-24 19:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-24 19:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-24 19:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-24 19:00
Moonshot AI eyes $50B valuation, Hong Kong IPO after Kimi K3 breakthrough
Reddit · 2026-07-24 19:00
I've been chasing the question of what algorithms a transformer can actually express -- separate from what it can learn. So I built a compiler: define a computa
Reddit · 2026-07-24 19:00
Built an open-source AI coding agent that was 7%–75% cheaper than a cold "claude -p" run on 6/6 well-localized tasks across repositories up to ~82k LOC. The big
HuggingFace Papers · 2026-07-24 07:00
Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candidate can o
HuggingFace Papers · 2026-07-24 07:00
Reinforcement learning for large language models (LLMs) typically relies on trust-region masks to stabilize off-policy updates. The dominant PPO-style approach
HuggingFace Papers · 2026-07-24 07:00
Agentic Reasoning has become a transformative force in financial analysis due to its ability to integrate large-scale information and generate reliable and accu
HuggingFace Papers · 2026-07-24 07:00
Real-world agent learning is often constrained by costly environment interactions, such as running time-consuming experiments or obtaining human feedback. In-co
HuggingFace Papers · 2026-07-24 07:00
Traditional agent development is split across prompt templates, tool schemas, callback code, and workflow graphs. We present NVIDIA Object-Oriented Agents (NOOA
HuggingFace Papers · 2026-07-24 07:00
As LLMs become more capable, they are increasingly deployed as collaborative agents, taking on user-delegated tasks through iterative interaction. Yet genuine i
HuggingFace Papers · 2026-07-24 07:00
Spatial intelligence is essential for agents to move from static semantic understanding toward interacting with the physical world. Many spatial tasks are groun
HuggingFace Papers · 2026-07-24 07:00
We introduce Tencent WorkBuddy Bench, a multi-domain evaluation suite for coding agents; this report documents its construction methodology, scoring protocol, a
HuggingFace Papers · 2026-07-24 07:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-24 07:00
Large language models are increasingly used in K-12 education, but existing benchmarks mainly test exam question answering rather than understanding how curricu
HuggingFace Papers · 2026-07-24 07:00
Embodied visual tracking (EVT) requires a mobile agent to continuously follow a specific target described in natural language using only onboard vision. While r
HuggingFace Papers · 2026-07-24 07:00
Deploying navigation systems at scale requires a recipe that minimizes sensor assumptions, generalizes across robot embodiments, and trains efficiently. Yet, to
HuggingFace Papers · 2026-07-24 07:00
Understanding motion in video is a fundamental challenge for visual learning, as frame-to-frame change entangles two sources of dynamics: camera motion and obje
HuggingFace Papers · 2026-07-24 07:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-24 07:00
The recent emergence of vibe-coding workflows is changing what coding agents are expected to do. Instead of merely completing code under fully specified instruc
HuggingFace Papers · 2026-07-24 07:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-24 07:00
LLM agent failures are difficult to debug because the step where an error surfaces is often not the one that caused it. Existing observability tools replay exec
HuggingFace Papers · 2026-07-24 07:00
The development of generalizable robotic manipulation policies is inherently bounded by the availability of large-scale, high-fidelity scene data. While recent
HuggingFace Papers · 2026-07-24 07:00
On-policy self-distillation (OPSD) is promising as it removes the external teacher required by on-policy distillation (OPD), yet it still needs asymmetric infor
HuggingFace Papers · 2026-07-24 07:00
Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evolve acro
HuggingFace Papers · 2026-07-24 07:00
We study sinusoidal recurrence as an iterative mechanism for harmonic spectral enrichment in implicit neural representations (INRs). Our analysis reveals that s
HuggingFace Papers · 2026-07-24 07:00
Controllable video generation remains challenging due to the difficulty of specifying precise multi-object interactions using text prompts or motion-control inp
HuggingFace Papers · 2026-07-24 07:00
We introduce SANA-Video 2.0, a hybrid video diffusion transformer instantiated at 5B and 14B scales under a unified architecture. Designed to generate high-qual
HuggingFace Papers · 2026-07-24 07:00
When a real-world scene is captured by a smartphone camera and viewed on its screen, the displayed image often differs noticeably from the original scene in col
OpenReview · 2026-07-24 07:00
The advancements in Large Language Models (LLMs) have been hindered by their substantial sizes, which necessitate LLM compression methods for practical deployme
OpenReview · 2026-07-24 07:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
GitHub Trending · 2026-07-24 07:00
🔥🔥🔥 AI-driven database tool and SQL client, The hottest GUI client, supporting MySQL, Oracle, PostgreSQL, DB2, SQL Server, DB2, SQLite, H2, ClickHouse, and more
GitHub Trending · 2026-07-24 07:00
《动手学大模型Dive into LLMs》系列编程实践教程
GitHub Trending · 2026-07-24 07:00
Pretty fancy and modern terminal file manager
Lobste.rs · 2026-07-24 07:00
If you train or serve models, you depend on MLIR whether or not you have ever written a line of it. XLA lowers through it, Triton is built on it, Mojo is MLIR-n
Lobste.rs · 2026-07-24 07:00
Article on using linear types for “naked” existential type variables.
Google Scholar · 2026-07-24 07:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-24 07:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-24 07:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-24 07:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-24 07:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-24 07:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Reddit · 2026-07-24 07:00
Its been almost 24 hours since reviews were released and I dont see the meta review still. Some people on reddit are saying they can see it. NeurIPS website say
Reddit · 2026-07-24 07:00
Hi, I'm new with ACM conferences. I have 2 papers at workshops and the conference website says: "Each workshop paper needs to be associated with one workshop-on
HuggingFace Papers · 2026-07-23 19:00
Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byproduct compounded
HuggingFace Papers · 2026-07-23 19:00
Evaluating the factuality of long-form generations has focused predominantly on precision, measuring whether the claims a model makes are correct. The dominant
HuggingFace Papers · 2026-07-23 19:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-23 19:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-23 19:00
Contextual entrainment is the tendency of a model to let auxiliary context in its input pull its output, independently of whether that context is relevant, true
HuggingFace Papers · 2026-07-23 19:00
LLM agent failures are difficult to debug because the step where an error surfaces is often not the one that caused it. Existing observability tools replay exec
HuggingFace Papers · 2026-07-23 19:00
Real-time EEG classification on edge devices is bottlenecked by the floating-point arithmetic of conventional neural networks. We investigated Differentiable Lo
HuggingFace Papers · 2026-07-23 19:00
Text-to-video generation has advanced significantly over the past five years through scaling of model size, data, and compute. Unlike model architecture, traini
HuggingFace Papers · 2026-07-23 19:00
Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale generative stack f
HuggingFace Papers · 2026-07-23 19:00
Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex prompts that
HuggingFace Papers · 2026-07-23 19:00
Simultaneous localization and mapping (SLAM) is one of the fundamental problems in robotics, as it enables autonomous operations in real-world scenarios. Under
OpenReview · 2026-07-23 19:00
The advancements in Large Language Models (LLMs) have been hindered by their substantial sizes, which necessitate LLM compression methods for practical deployme
OpenReview · 2026-07-23 19:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
Lobste.rs · 2026-07-23 19:00
Breaking down the Hugging Face security incident caused by OpenAI's own models during a benchmark run - the sandbox escape, the package proxy, and whether it's
Lobste.rs · 2026-07-23 19:00
Not just development, distribution of software may change as well
Lobste.rs · 2026-07-23 19:00
Exploit development for the Windows WalletService vulnerability, from caller-controlled known-folder resolution to a persisted ESE callback and an interactive S
Lobste.rs · 2026-07-23 19:00
A vulnerability in macOS allows an attacker to silently replace the main executable of any application downloaded from the web without requiring elevated privil
Google Scholar · 2026-07-23 19:00
This paper conducts systematic analysis of advancements in the field of artificial intelligence (AI) from the year 2010 onwards, by performing original quantita
Google Scholar · 2026-07-23 19:00
Large language models (LLM) in computational social science: prospects, current state, and challenges
Google Scholar · 2026-07-23 19:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-23 19:00
Quantum machine learning: A comprehensive review of integrating AI with quantum computing for computational advancements
Google Scholar · 2026-07-23 19:00
Artificial intelligence and machine learning in cybersecurity: a deep dive into state-of-the-art techniques and future paradigms
Google Scholar · 2026-07-23 19:00
Catalyst breakthroughs in methane dry reforming: Employing machine learning for future advancements
Hacker News · 2026-07-23 19:00
How the most-used zero-knowledge proof system (Groth16) works
Hacker News · 2026-07-23 19:00
ZK/SEC Quarterly Back to all posts Archetype x zkSecurity - Proof is in the Pudding: Groth16 ZK/SEC July 23, 2026 2 min read educative zk groth16 For the 10th s
Hacker News · 2026-07-23 19:00
Show HN: Mumble Dictation – local dictation that learns your vocabulary
Hacker News · 2026-07-23 19:00
Google announces Gemini 3.6 Flash and cybersecurity AI, teases 3.5 Pro, Gemini 4
Reddit · 2026-07-23 19:00
I have been working on an MCP workflow for implementing deep learning models from an engineering plan. This is useful for ml engineers etc. who want a more stru
Reddit · 2026-07-23 19:00
Hi everyone, I recently received an interview offer for ML position at Adyen, and the first round will be a live coding round on HackerRank. I scheduled it for
Reddit · 2026-07-23 19:00
The interesting finding from a new [arXiv paper]( https://arxiv.org/abs/2607.16165 ) isn't that a frontier vision model failed a new benchmark, that happens wee
Reddit · 2026-07-23 19:00
good luck! submitted by /u/Business-Kale-1406 [link] [comments]
Reddit · 2026-07-23 19:00
NeurIPS E and D track review are out today and the average rating I received is a 3 and confidence is a 4. I can correct and address all their concerns. Do I st
Reddit · 2026-07-23 19:00
The reviews were just released, and I downloaded my paper from OpenReview to identify areas that needed improvement. However, GPT warned me that the PDF contain
HuggingFace Papers · 2026-07-23 07:00
As autonomous agents rapidly evolve, their ability to reliably manipulate ubiquitous digital documents has become critical for enabling general-purpose AI assis
HuggingFace Papers · 2026-07-23 07:00
Asynchronous reinforcement learning improves throughput by decoupling rollout generation from optimization, but staleness is an inevitable byproduct compounded
HuggingFace Papers · 2026-07-23 07:00
Large language models can answer scientific questions, yet a correct output does not reveal whether the model represents or uses the governing physics. Here we
HuggingFace Papers · 2026-07-23 07:00
Injecting factual knowledge into large language models (LLMs) reliably and at scale remains an open challenge. Hypernetworks provide a promising solution to lar
HuggingFace Papers · 2026-07-23 07:00
Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, includi
HuggingFace Papers · 2026-07-23 07:00
Reinforcement learning (RL) has become a dominant paradigm for enhancing LLMs' reasoning capabilities. However, RL algorithms with PPO-Clip are inherently limit
HuggingFace Papers · 2026-07-23 07:00
Reinforcement learning with verifiable rewards has become the predominant recipe for eliciting test-time scaling in explicit Chain-of-Thought reasoners. Yet thi
HuggingFace Papers · 2026-07-23 07:00
Evaluating the factuality of long-form generations has focused predominantly on precision, measuring whether the claims a model makes are correct. The dominant
HuggingFace Papers · 2026-07-23 07:00
Human vision is a closed loop: gaze is continuously redirected by intermediate hypotheses rather than a single snapshot. Decades of psychophysics and cognitive
HuggingFace Papers · 2026-07-23 07:00
Finetuning a pretrained vision-language model (VLM) on robot demonstrations via behavior cloning (BC) has become the standard recipe for vision-language-action
HuggingFace Papers · 2026-07-23 07:00
Reinforcement learning with verifiable rewards (RLVR) has substantially improved language-model reasoning, yet its extension to vision-language models remains c
HuggingFace Papers · 2026-07-23 07:00
Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training: on a 6.78B-parameter MoE language model, AdamW keeps
HuggingFace Papers · 2026-07-23 07:00
As large language models and AI agents become the primary consumers of search results, document set quality determines the upper bound of downstream generation.
HuggingFace Papers · 2026-07-23 07:00
Modern ASR models trained on heterogeneously annotated data treat transcription style (verbatim vs. intended) as an uncontrolled latent variable, causing measur
HuggingFace Papers · 2026-07-23 07:00
Practical robotic grasping in complex scenes requires both 3D spatial reasoning and alignment with task-specific requirements. Vision-language models (VLMs) off
HuggingFace Papers · 2026-07-23 07:00
LLM agent failures are difficult to debug because the step where an error surfaces is often not the one that caused it. Existing observability tools replay exec
HuggingFace Papers · 2026-07-23 07:00
Recent autoregressive video diffusion methods are increasingly built upon Self Forcing, where the student is trained on histories produced by its own rollout ra
HuggingFace Papers · 2026-07-23 07:00
Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale generative stack f
HuggingFace Papers · 2026-07-23 07:00
Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex prompts that
HuggingFace Papers · 2026-07-23 07:00
This work introduces G-MAD, an open-source framework that uses Arma3 to generate synchronized multi-view RGB-T data for aerial object detection. G-MAD addresses
HuggingFace Papers · 2026-07-23 07:00
Video Diffusion Transformers process long spatio-temporal sequences, making self-attention the main bottleneck in high-resolution video generation. Training-fre
OpenReview · 2026-07-23 07:00
The advancements in Large Language Models (LLMs) have been hindered by their substantial sizes, which necessitate LLM compression methods for practical deployme
OpenReview · 2026-07-23 07:00
Recent advancements in 3D generation are predominantly propelled by improvements in 3D-aware image diffusion models. These models are pretrained on Internet-sca
GitHub Trending · 2026-07-23 07:00
Offline, privacy-first grammar checker. Fast, open-source, Rust-powered
GitHub Trending · 2026-07-23 07:00
Open-source & free — Battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, b
GitHub Trending · 2026-07-23 07:00
The best browser for both you and your AI agents work in parallel.
GitHub Trending · 2026-07-23 07:00
The Free Software Media System - Server Backend & API
GitHub Trending · 2026-07-23 07:00
Trending repository Julian-adv/OpenMMO.
Lobste.rs · 2026-07-23 07:00
Appaji, a Computer Science graduate from IIT Patna and former intern at Arista Networks, is a Software Engineer at Infinite Reality. Passionate about building i
Lobste.rs · 2026-07-23 07:00
Richard Hamming famously used to ask his colleagues at Bell Labs this question: “What is the most important problem in your field, and why aren’t you working on
Google Scholar · 2026-07-23 07:00
AI alignment is a human problem
Google Scholar · 2026-07-23 07:00
This work is licensed under a Creative Commons Attribution 4.0 International License .
Google Scholar · 2026-07-23 07:00
As AI capabilities advance toward and potentially beyond human-level performance, a natural transition emerges where AI-driven development becomes more efficien
Google Scholar · 2026-07-23 07:00
Cluster Computing - With the emergence of quantum computers, traditional cryptographic methods are vulnerable to attacks, emphasizing the need for post-quantum
Google Scholar · 2026-07-23 07:00
The progression of artificial intelligence (AI) technologies has reached a level that greatly enhances the different organizational sectors by facilitating them
Google Scholar · 2026-07-23 07:00
Cybersecurity is now a major issue, and it is getting the attention of researchers, academics, and businesses to really focus on safeguarding our information sy
Google Scholar · 2026-07-23 07:00
Explore millions of resources from scholarly journals, books, newspapers, videos and more, on the ProQuest Platform.
Google Scholar · 2026-07-23 07:00
Utilisation of Artificial Intelligence and Cybersecurity Capabilities: A Symbiotic Relationship for Enhanced Security and Applicability
Google Scholar · 2026-07-23 07:00
Generative AI-enhanced cybersecurity framework for enterprise data privacy management
Google Scholar · 2026-07-23 07:00
Machine learning approach for mapping the heat capacity of deep eutectic solvents for sustainable energy applications
Google Scholar · 2026-07-23 07:00
Construction payment automation through scan-to-BIM and blockchain-enabled smart contract
Hacker News · 2026-07-23 07:00
As the digital landscape evolves at an unprecedented pace, mastering the latest technological shifts is critical for platform growth. In this comprehensive AI
HuggingFace Papers · 2026-07-21 19:00
We present S1-Omni, a unified multimodal reasoning model for scientific understanding, prediction, and generation. AI for Science (AI4S) has advanced significan
HuggingFace Papers · 2026-07-21 19:00
Training tool-use agents to improve from their own experience remains challenging, as supervised fine-tuning relies on fixed teacher-distilled trajectories, whi
HuggingFace Papers · 2026-07-21 19:00
Reinforcement learning from verifiable rewards (e.g. GRPO) is the engine behind today's reasoning models, yet it grades only the final answer. On hard problems
HuggingFace Papers · 2026-07-21 19:00
Autonomous negotiation agents are increasingly deployed in high-stakes settings such as insurance and procurement. While cryptographic techniques protect explic
HuggingFace Papers · 2026-07-21 19:00
Video multimodal large language models (MLLMs) can describe what happens in a video, but rarely identify when the supporting evidence occurs. We study generalis
HuggingFace Papers · 2026-07-21 19:00
Modern video generation models are increasingly hailed as emerging world models with an internalized grasp of physical law. Yet existing benchmarks largely eval
HuggingFace Papers · 2026-07-21 19:00
Predicting a football match before kickoff requires more than knowing past results: a model must use changing information and make a clear prediction before the
HuggingFace Papers · 2026-07-21 19:00
Multi-agent systems routinely place one AI agent in authority over another. When a subordinate refuses a task, the manager chooses the outcome: it can renegotia
HuggingFace Papers · 2026-07-21 19:00
Despite strong capabilities in data understanding and decision-making, autonomous data science agents still heavily rely on trial-and-error workflows that invol
HuggingFace Papers · 2026-07-21 19:00
Large language model (LLM) post-training is essential for improving reasoning, adaptation, and alignment. Existing methods mainly follow two paradigms: reinforc
HuggingFace Papers · 2026-07-21 19:00
Recent growth in reinforcement learning (RL) has surfaced a need for diverse, specialized training environments. Hand-curated environments with fixed task and r
HuggingFace Papers · 2026-07-21 19:00
Healthcare spans high-stakes communication, expert reasoning, and workflow execution, yet specialized LLMs that cover these use cases together remain limited. A
HuggingFace Papers · 2026-07-21 19:00
Code review helps maintain software quality before code integration, but it also imposes a substantial workload on human reviewers. As generative artificial int
HuggingFace Papers · 2026-07-21 19:00
Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, exploit development
HuggingFace Papers · 2026-07-21 19:00
We introduce Self-Verified Reasoner (SVR-R1), a multi-turn RL framework that turns a model's own verification into a learning signal for multimodal reasoning. F
HuggingFace Papers · 2026-07-21 19:00
Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent
HuggingFace Papers · 2026-07-21 19:00
Training API-calling large language model (LLM) agents demands massive amounts of high-quality trajectories. However, collecting such data at scale typically re
HuggingFace Papers · 2026-07-21 19:00
We propose Token-Level Off-Policy Labeling (TOPL), an off-policy training paradigm that reframes post-training as a token-level correctness prediction task. Our
HuggingFace Papers · 2026-07-21 19:00
The post-training of Vision-Language-Action (VLA) models is essential due to the diversity of simulators, robot embodiments, and task objectives. Existing compu
HuggingFace Papers · 2026-07-21 19:00
Entropy control has become an effective tool in reinforcement learning (RL) of large language models (LLMs), helping balance exploration-exploitation trade-off
HuggingFace Papers · 2026-07-21 19:00
We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a wide range o
HuggingFace Papers · 2026-07-21 19:00
We present Audio-Visual Flamingo (AV-Flamingo), a fully open state-of-the-art audio-visual large language model (AV-LLM) for joint understanding and reasoning o
HuggingFace Papers · 2026-07-21 19:00
Reinforcement learning with verifiable rewards (RLVR) commonly uses entropy for advantage shaping. However, entropy cannot distinguish useful uncertainty from d
HuggingFace Papers · 2026-07-21 19:00
Agentic language models must learn when to call tools, when to consume tool responses, and when to answer directly. This makes multi-teacher on-policy distillat
HuggingFace Papers · 2026-07-21 19:00
Building assistants that can continually watch the world, remember what they see, and reason over their accumulated experience is a long-standing goal, and rece
HuggingFace Papers · 2026-07-21 19:00
Despite recent scaling successes, multilingual ASR performance remains highly uneven, with long-tail languages suffering from severe data scarcity. This work ad
HuggingFace Papers · 2026-07-21 19:00
The prevailing inference framework for diffusion models formulates generation fundamentally as a problem of numerical integration. This perspective casts the mo
HuggingFace Papers · 2026-07-21 19:00
Temporal grounding in long recordings remains challenging for audio-conditioned LLMs. We present a time-aware audio LLM that answers questions with explicit tim
HuggingFace Papers · 2026-07-21 19:00
Optical coherence tomography (OCT) imaging is essential for the diagnosis and treatment of retinal diseases. Although multimodal large language models (MLLMs) h
HuggingFace Papers · 2026-07-21 19:00
Under model--harness co-evolution, harnesses are not merely inference-time scaffolds but data-generating components whose execution traces can shape future foun
HuggingFace Papers · 2026-07-21 19:00
Egocentric videos of human manipulation provide scalable supervision for embodied intelligence, yet existing resources rarely combine low-cost continuous captur
HuggingFace Papers · 2026-07-21 19:00
Vision-language-action (VLA) models predict robot actions from visual observations and language instructions. These actions are defined in the robot's own 3D co
HuggingFace Papers · 2026-07-21 19:00
We present RynnBrain 1.1, a family of embodied foundation models spanning 2B, 9B, and 122B-A10B scales. Trained with a unified spatio-temporal and physically gr
HuggingFace Papers · 2026-07-21 19:00
Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill libraries are mos
HuggingFace Papers · 2026-07-21 19:00
Plasma diagnostic models for tokamak fusion devices are almost universally evaluated on clean, complete sensor data. In practice, fusion diagnostics fail regula
HuggingFace Papers · 2026-07-21 19:00
Graph retrieval-augmented generation (GraphRAG) enhances large language models with structured knowledge, yet existing systems construct knowledge graphs in a s
HuggingFace Papers · 2026-07-21 19:00
Reinforcement learning (RL) on open-ended tasks compresses an LLM's rubric-based evaluation into a scalar reward, discarding rich textual feedback and conflatin
HuggingFace Papers · 2026-07-21 19:00
Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if k verifier calls all accept it. Under condit
HuggingFace Papers · 2026-07-21 19:00
Scaling robust driving policies is fundamentally bottlenecked by the scarcity of edge cases in curated datasets. While the real world continuously captures thes
HuggingFace Papers · 2026-07-21 19:00
We present AffectFlow-DINO, a multi-task learning system for the 11th ABAW challenge that extends a standard deterministic architecture with a conditional recti
HuggingFace Papers · 2026-07-21 19:00
On-policy distillation is an alternative post-training method in reinforcement learning that alleviates the constraints imposed by reward models by providing to
HuggingFace Papers · 2026-07-21 19:00
In this report, we introduce Qwen-Music, a powerful music generation model capable of producing highly musical and high-fidelity songs with complete vocal singi
HuggingFace Papers · 2026-07-21 19:00
We present Loopie, the most powerful looped Transformer to date. The Loopie series consists of two Mixture-of-Experts (MoE) models: a 20B-parameter model with 2
HuggingFace Papers · 2026-07-21 19:00
Video generative models commonly rely on latent spaces learned by 3D Variational Autoencoders (3D-VAEs). However, conventional 3D-VAEs are mainly optimized for
HuggingFace Papers · 2026-07-21 07:00
We present S1-Omni, a unified multimodal reasoning model for scientific understanding, prediction, and generation. AI for Science (AI4S) has advanced significan
HuggingFace Papers · 2026-07-21 07:00
Training tool-use agents to improve from their own experience remains challenging, as supervised fine-tuning relies on fixed teacher-distilled trajectories, whi
HuggingFace Papers · 2026-07-21 07:00
Reinforcement learning from verifiable rewards (e.g. GRPO) is the engine behind today's reasoning models, yet it grades only the final answer. On hard problems
HuggingFace Papers · 2026-07-21 07:00
Autonomous negotiation agents are increasingly deployed in high-stakes settings such as insurance and procurement. While cryptographic techniques protect explic
HuggingFace Papers · 2026-07-21 07:00
Video multimodal large language models (MLLMs) can describe what happens in a video, but rarely identify when the supporting evidence occurs. We study generalis
HuggingFace Papers · 2026-07-21 07:00
Modern video generation models are increasingly hailed as emerging world models with an internalized grasp of physical law. Yet existing benchmarks largely eval
HuggingFace Papers · 2026-07-21 07:00
Predicting a football match before kickoff requires more than knowing past results: a model must use changing information and make a clear prediction before the
HuggingFace Papers · 2026-07-21 07:00
Despite strong capabilities in data understanding and decision-making, autonomous data science agents still heavily rely on trial-and-error workflows that invol
HuggingFace Papers · 2026-07-21 07:00
Large language model (LLM) post-training is essential for improving reasoning, adaptation, and alignment. Existing methods mainly follow two paradigms: reinforc
HuggingFace Papers · 2026-07-21 07:00
Healthcare spans high-stakes communication, expert reasoning, and workflow execution, yet specialized LLMs that cover these use cases together remain limited. A
HuggingFace Papers · 2026-07-21 07:00
Code review helps maintain software quality before code integration, but it also imposes a substantial workload on human reviewers. As generative artificial int
HuggingFace Papers · 2026-07-21 07:00
Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, exploit development
HuggingFace Papers · 2026-07-21 07:00
We introduce Self-Verified Reasoner (SVR-R1), a multi-turn RL framework that turns a model's own verification into a learning signal for multimodal reasoning. F
HuggingFace Papers · 2026-07-21 07:00
Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent
HuggingFace Papers · 2026-07-21 07:00
Training API-calling large language model (LLM) agents demands massive amounts of high-quality trajectories. However, collecting such data at scale typically re
HuggingFace Papers · 2026-07-21 07:00
We propose Token-Level Off-Policy Labeling (TOPL), an off-policy training paradigm that reframes post-training as a token-level correctness prediction task. Our
HuggingFace Papers · 2026-07-21 07:00
The post-training of Vision-Language-Action (VLA) models is essential due to the diversity of simulators, robot embodiments, and task objectives. Existing compu
HuggingFace Papers · 2026-07-21 07:00
Entropy control has become an effective tool in reinforcement learning (RL) of large language models (LLMs), helping balance exploration-exploitation trade-off
HuggingFace Papers · 2026-07-21 07:00
We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a wide range o
HuggingFace Papers · 2026-07-21 07:00
We present Audio-Visual Flamingo (AV-Flamingo), a fully open state-of-the-art audio-visual large language model (AV-LLM) for joint understanding and reasoning o
HuggingFace Papers · 2026-07-21 07:00
Reinforcement learning with verifiable rewards (RLVR) commonly uses entropy for advantage shaping. However, entropy cannot distinguish useful uncertainty from d
HuggingFace Papers · 2026-07-21 07:00
Building assistants that can continually watch the world, remember what they see, and reason over their accumulated experience is a long-standing goal, and rece
HuggingFace Papers · 2026-07-21 07:00
Despite recent scaling successes, multilingual ASR performance remains highly uneven, with long-tail languages suffering from severe data scarcity. This work ad
HuggingFace Papers · 2026-07-21 07:00
The prevailing inference framework for diffusion models formulates generation fundamentally as a problem of numerical integration. This perspective casts the mo
HuggingFace Papers · 2026-07-21 07:00
Temporal grounding in long recordings remains challenging for audio-conditioned LLMs. We present a time-aware audio LLM that answers questions with explicit tim
HuggingFace Papers · 2026-07-21 07:00
Optical coherence tomography (OCT) imaging is essential for the diagnosis and treatment of retinal diseases. Although multimodal large language models (MLLMs) h
HuggingFace Papers · 2026-07-21 07:00
Under model--harness co-evolution, harnesses are not merely inference-time scaffolds but data-generating components whose execution traces can shape future foun
HuggingFace Papers · 2026-07-21 07:00
Vision-language-action (VLA) models predict robot actions from visual observations and language instructions. These actions are defined in the robot's own 3D co
HuggingFace Papers · 2026-07-21 07:00
We present RynnBrain 1.1, a family of embodied foundation models spanning 2B, 9B, and 122B-A10B scales. Trained with a unified spatio-temporal and physically gr
HuggingFace Papers · 2026-07-21 07:00
Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill libraries are mos
HuggingFace Papers · 2026-07-21 07:00
Plasma diagnostic models for tokamak fusion devices are almost universally evaluated on clean, complete sensor data. In practice, fusion diagnostics fail regula
HuggingFace Papers · 2026-07-21 07:00
Graph retrieval-augmented generation (GraphRAG) enhances large language models with structured knowledge, yet existing systems construct knowledge graphs in a s
HuggingFace Papers · 2026-07-21 07:00
Reinforcement learning (RL) on open-ended tasks compresses an LLM's rubric-based evaluation into a scalar reward, discarding rich textual feedback and conflatin
HuggingFace Papers · 2026-07-21 07:00
Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if k verifier calls all accept it. Under condit
HuggingFace Papers · 2026-07-21 07:00
Scaling robust driving policies is fundamentally bottlenecked by the scarcity of edge cases in curated datasets. While the real world continuously captures thes
HuggingFace Papers · 2026-07-21 07:00
We present AffectFlow-DINO, a multi-task learning system for the 11th ABAW challenge that extends a standard deterministic architecture with a conditional recti
HuggingFace Papers · 2026-07-21 07:00
On-policy distillation is an alternative post-training method in reinforcement learning that alleviates the constraints imposed by reward models by providing to
HuggingFace Papers · 2026-07-21 07:00
In this report, we introduce Qwen-Music, a powerful music generation model capable of producing highly musical and high-fidelity songs with complete vocal singi
HuggingFace Papers · 2026-07-21 07:00
We present Loopie, the most powerful looped Transformer to date. The Loopie series consists of two Mixture-of-Experts (MoE) models: a 20B-parameter model with 2
HuggingFace Papers · 2026-07-21 07:00
Video generative models commonly rely on latent spaces learned by 3D Variational Autoencoders (3D-VAEs). However, conventional 3D-VAEs are mainly optimized for
HuggingFace Papers · 2026-07-21 07:00
Pruning long context for coding agents has been a vital technology for efficient context management. While existing context pruning methods such as SWE-Pruner r
HuggingFace Papers · 2026-07-21 07:00
Hyper-Connections (HC) expand the residual stream of Transformers into N parallel streams, providing a form of memory scaling beyond model width and depth. Mani
HuggingFace Papers · 2026-07-21 07:00
We present a continuous geometric framework that models the discrete algebraic operations of the Transformer architecture as an integro-differential equation (I
HuggingFace Papers · 2026-07-21 07:00
Current video generation models achieve impressive results in single-shot generation, yet remain limited in cinematic video generation, where coherent narrative
HuggingFace Papers · 2026-07-20 19:00
We present S1-Omni, a unified multimodal reasoning model for scientific understanding, prediction, and generation. AI for Science (AI4S) has advanced significan
HuggingFace Papers · 2026-07-20 19:00
Reinforcement learning from verifiable rewards (e.g. GRPO) is the engine behind today's reasoning models, yet it grades only the final answer. On hard problems
HuggingFace Papers · 2026-07-20 19:00
Autonomous negotiation agents are increasingly deployed in high-stakes settings such as insurance and procurement. While cryptographic techniques protect explic
HuggingFace Papers · 2026-07-20 19:00
Despite strong capabilities in data understanding and decision-making, autonomous data science agents still heavily rely on trial-and-error workflows that invol
HuggingFace Papers · 2026-07-20 19:00
Healthcare spans high-stakes communication, expert reasoning, and workflow execution, yet specialized LLMs that cover these use cases together remain limited. A
HuggingFace Papers · 2026-07-20 19:00
Code review helps maintain software quality before code integration, but it also imposes a substantial workload on human reviewers. As generative artificial int
HuggingFace Papers · 2026-07-20 19:00
Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, exploit development
HuggingFace Papers · 2026-07-20 19:00
We introduce Self-Verified Reasoner (SVR-R1), a multi-turn RL framework that turns a model's own verification into a learning signal for multimodal reasoning. F
HuggingFace Papers · 2026-07-20 19:00
Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent
HuggingFace Papers · 2026-07-20 19:00
We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a wide range o
HuggingFace Papers · 2026-07-20 19:00
We present Audio-Visual Flamingo (AV-Flamingo), a fully open state-of-the-art audio-visual large language model (AV-LLM) for joint understanding and reasoning o
HuggingFace Papers · 2026-07-20 19:00
Reinforcement learning with verifiable rewards (RLVR) commonly uses entropy for advantage shaping. However, entropy cannot distinguish useful uncertainty from d
HuggingFace Papers · 2026-07-20 19:00
Under model--harness co-evolution, harnesses are not merely inference-time scaffolds but data-generating components whose execution traces can shape future foun
HuggingFace Papers · 2026-07-20 19:00
Vision-language-action (VLA) models predict robot actions from visual observations and language instructions. These actions are defined in the robot's own 3D co
HuggingFace Papers · 2026-07-20 19:00
Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill libraries are mos
HuggingFace Papers · 2026-07-20 19:00
Plasma diagnostic models for tokamak fusion devices are almost universally evaluated on clean, complete sensor data. In practice, fusion diagnostics fail regula