AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,888 results
11 May 2026

SWaRL: Safeguard Code Watermarking via Reinforcement Learning

ResearchDGX agent

arXiv:2601.02602v2 Announce Type: replace-cross Abstract: We present SWaRL, a robust and fidelity-preserving watermarking framework designed to protect the intellectual property of code LLMs by embedd

Synergistic Benefits of Joint Molecule Generation and Property Prediction

ResearchDGX agent

arXiv:2504.16559v3 Announce Type: replace Abstract: Modeling the joint distribution of data samples and their properties allows to construct a single model for both data generation and property predic

Task-Oriented Communication for Human Action Understanding via Edge-Cloud Co-Inference

ResearchDGX agent

arXiv:2605.07354v1 Announce Type: cross Abstract: The expanding application of smart sensing has created a growing demand for the accurate understanding of human action at the network edge. Traditiona

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TCMIIES: A Browser-Based LLM-Powered Intelligent Information Extraction System for Academic Literature

Local AiDGX agent

arXiv:2605.07507v1 Announce Type: new Abstract: The exponential growth of academic publications has created an urgent need for automated tools capable of extracting structured knowledge from unstructu

Teacher-Feature Drifting: One-Step Diffusion Distillation with Pretrained Diffusion Representations

ResearchDGX agent

arXiv:2605.07327v1 Announce Type: new Abstract: Sampling from pretrained diffusion and flow-matching models typically requires many forward passes to generate diverse and high-fidelity images. Existin

Tessellations of Semi-Discrete Flow Matching

ResearchDGX agent

arXiv:2605.07513v1 Announce Type: new Abstract: We study Flow Matching in a semi-discrete setting where a Gaussian source is transported toward a discrete target supported on finitely many points. Thi

The Download: the hantavirus outbreak and Musk v. Altman week 2

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Here’s what you need to know about the cruise ship hantavirus

The Effective Depth Paradox: Evaluating the Relationship between Architectural Topology and Trainability in Deep CNNs

ResearchDGX agent

arXiv:2602.13298v3 Announce Type: replace-cross Abstract: This paper investigates the relationship between convolutional neural network (CNN) topology and image recognition performance through a compa

The Minimax Rate of Second-Order Calibration

ResearchDGX agent

arXiv:2605.07808v1 Announce Type: new Abstract: We characterize the minimax rate of estimating the second-order calibration error for binary classification, which quantifies whether a higher-order pre

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters …

ResearchDGX agent

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters in the new world We are at Step 1. Interaction Models are gr

This works really well btw, at the end of your query ask your LLM to 'structure your response as HTML', then view the generated file in your…

ResearchDGX agent

This works really well btw, at the end of your query ask your LLM to 'structure your response as HTML', then view the generated file in your browser. I've also had some success asking the LLM to prese

Three-in-One World Model: Energy-Based Consistency, Prediction, and Counterfactual Inference for Marketing Intervention

ResearchDGX agent

arXiv:2605.07199v1 Announce Type: new Abstract: Marketing decisions reflect the interaction of latent consumer heterogeneity, time-varying internal states, and explicit interventions, a structure that

TimeLesSeg: Unified Contrast-Agnostic Cross-Sectional and Longitudinal MS Lesion Segmentation via a Stochastic Generative Model

ResearchDGX agent

arXiv:2605.07955v1 Announce Type: cross Abstract: Multiple sclerosis (MS) expresses substantial clinical and radiological heterogeneity, which poses significant challenges for automatic lesion segment

To be clear, I meant 'idea' in the context of science and engineering, in particular AI, as in 'theory' or 'invention', like a new theory of…

ResearchDGX agent

To be clear, I meant 'idea' in the context of science and engineering, in particular AI, as in 'theory' or 'invention', like a new theory of cognition or an idea for a new ML approach. I thought that

Today we're sharing our work on interaction models. A new class of model trained from scratch to handle real-time interaction natively, inst…

ResearchDGX agent

Today we're sharing our work on interaction models. A new class of model trained from scratch to handle real-time interaction natively, instead of gluing it onto a turn-based one. https://youtu.be/A12

TopoPrune: Robust Data Pruning via Unified Latent Space Topology

ResearchDGX agent

arXiv:2602.02739v2 Announce Type: replace-cross Abstract: Geometric data pruning methods, while practical for leveraging pretrained models, are fundamentally unstable. Their reliance on extrinsic geom

Toward Privileged Foundation Models:LUPI for Accelerated and Improved Learning

ResearchDGX agent

arXiv:2605.07799v1 Announce Type: cross Abstract: Training foundation models is computationally intensive and often slow to converge.We introduce PIQL,Privileged Information for Quick and Quality Lear

Towards an Inferentialist Account of Information Through Proof-theoretic Semantics

ResearchDGX agent

arXiv:2605.05368v2 Announce Type: replace-cross Abstract: Information is one of the most widely-discussed concepts of the current era. However, a great deal of insightful work notwithstanding, it is y

Towards Explainable Industrial Anomaly Detection via Knowledge-Guided Latent Reasoning

ResearchDGX agent

arXiv:2602.09850v2 Announce Type: replace Abstract: Industrial anomaly detection demands precise reasoning over fine-grained defect patterns. However, existing multimodal large language models (MLLMs)

Towards Highly-Constrained Human Motion Generation with Retrieval-Guided Diffusion Noise Optimization

ResearchDGX agent

arXiv:2605.08054v1 Announce Type: new Abstract: Generating human motion that satisfies customized zero-shot goal functions, enabling applications such as controllable character animation and behavior

TRACE: Tourism Recommendation with Accountable Citation Evidence

ResearchDGX agent

arXiv:2605.07677v1 Announce Type: cross Abstract: Tourism is a high-stakes setting for conversational recommender systems (CRS): a plausible-sounding suggestion can waste real money and trip time once

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs

ResearchDGX agent

arXiv:2605.07568v1 Announce Type: cross Abstract: The Arrow-of-Time (AoT) task, determining whether a video plays forward or backward by recognizing temporal irreversibility, is one humans solve with

Transformer-Based Wildlife Species Classification from Daily Movement Trajectories

ResearchDGX agent

arXiv:2605.06726v1 Announce Type: new Abstract: Inferring the identity of wildlife species from daily movement data alone is a challenging task. We train sequence models on large-scale, 7-species GPS

TRAS: An Interactive Software for Tracing Tree Ring Cross Sections

ResearchDGX agent

arXiv:2605.08025v1 Announce Type: new Abstract: Tree ring marking remains a key step in dendrometry and dendrochronology, but it is often performed manually, making the process time-consuming, subject

TriP: A Triangle Puzzle Approach to Robust Translation Averaging

ResearchDGX agent

arXiv:2605.07143v1 Announce Type: new Abstract: Translation averaging aims to recover camera locations from pairwise relative translation directions and is a fundamental component of global Structure-

TTF: Temporal Token Fusion for Efficient Video-Language Model

ResearchDGX agent

arXiv:2605.07355v1 Announce Type: cross Abstract: Video-language models (VLMs) face rapid inference costs as visual token counts scale with video length. For example, 32 frames at 448{imes}448 resolut

Tyche: One Step Flow for Efficient Probabilistic Weather Forecasting

ResearchDGX agent

arXiv:2605.06916v1 Announce Type: new Abstract: Probabilistic weather forecasting requires not only accurate trajectories, but calibrated distributions over plausible atmospheric futures. Recent data-

Uncertainty-Aware Structured Data Extraction from Full CMR Reports via Distilled LLMs

ResearchDGX agent

arXiv:2605.08045v1 Announce Type: new Abstract: Converting free-text cardiac magnetic resonance (CMR) reports into auditable structured data remains a bottleneck for cohort assembly, longitudinal cura

Uncertainty Quantification for Cardiac Shape Reconstruction with Deep Signed Distance Functions via MCMC methods

ResearchDGX agent

arXiv:2605.07987v1 Announce Type: cross Abstract: Atlas-based approaches allow high-quality, patient-specific shape reconstructions of cardiac anatomy from sparse and/or noisy data such as point cloud

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions

ResearchDGX agent

arXiv:2605.07271v1 Announce Type: cross Abstract: Layer pruning efficiently reduces Large Language Model (LLM) computational costs but often triggers sudden performance collapse. Existing representati

UniISP: A Unified ISP Framework for Both Human and Machine Vision

ResearchDGX agent

arXiv:2605.07359v1 Announce Type: new Abstract: Compared to RGB images, raw sensor data provides a richer representation of information, which is crucial for accurate recognition, particularly under c

Upper Generalization Bounds for Neural Oscillators

ResearchDGX agent

arXiv:2603.09742v2 Announce Type: replace Abstract: Neural oscillators that originate from second-order ordinary differential equations (ODEs) have shown competitive performance in learning mappings b

VecCISC: Improving Confidence-Informed Self-Consistency with Reasoning Trace Clustering and Candidate Answer Selection

ResearchDGX agent

arXiv:2605.08070v1 Announce Type: new Abstract: A standard technique for scaling inference-time reasoning is Self-Consistency, whereby multiple candidate answers are sampled from an LLM and the most c

Velocity-Space 3D Asset Editing

ResearchDGX agent

arXiv:2605.07385v1 Announce Type: cross Abstract: Editing a 3D asset locally, modifying a target region while preserving the rest, is a fundamental requirement of native 3D editing. Existing methods e

VIMCAN: Visual-Inertial 3D Human Pose Estimation with Hybrid Mamba-Cross-Attention Network

ResearchDGX agent

arXiv:2605.07552v1 Announce Type: new Abstract: The rapid advances in deep learning have significantly enhanced the accuracy of multimodal 3D human pose estimation (HPE). However, the state-of-the-art

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing

ResearchDGX agent

arXiv:2605.06765v1 Announce Type: cross Abstract: Human speech conveys expressiveness beyond linguistic content, including personality, mood, or performance elements, such as a comforting tone or humm

Weather-Robust Scene Semantics with Vision-Aligned 4D Radar

ResearchDGX agent

arXiv:2605.07367v1 Announce Type: cross Abstract: Cameras and LiDAR degrade in rain, fog, and snow, while millimeter-wave radar remains largely unaffected. We align a radar encoder to frozen SigLIP vi

WeatherSyn: An Instruction Tuning MLLM For Weather Forecasting Report Generation

ResearchDGX agent

arXiv:2605.07522v1 Announce Type: new Abstract: Accurate weather forecast reporting enables individuals and communities to better plan daily activities and agricultural operations. However, the curren

What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion

ResearchDGX agent

arXiv:2605.07915v1 Announce Type: new Abstract: Tokenizers are a crucial component of latent diffusion models, as they define the latent space in which diffusion models operate. However, existing toke

What will we be like when he is gone? Can we return to mutual respect? Can we believe we are all on the same team as Obama and McCain did? C…

ResearchDGX agent

What will we be like when he is gone? Can we return to mutual respect? Can we believe we are all on the same team as Obama and McCain did? Can we imagine the mutual respect of those two, competitors b

When Diffusion Model Can Ignore Dimension: An Entropy-Based Theory

ResearchDGX agent

arXiv:2605.07969v1 Announce Type: new Abstract: Diffusion models perform remarkably well on high-dimensional data such as images, often using only a modest number of reverse-time steps. Despite this p

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

ResearchDGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

When Does Critique Improve AI-Assisted Theoretical Physics? SCALAR: Structured Critic--Actor Loop for Agentic Reasoning

Model ReleasesDGX agent

arXiv:2605.06772v1 Announce Type: new Abstract: As large language models (LLMs) show increasing promise on research-level physics reasoning tasks and agentic AI becomes more common, a practical questi

When Does Embedding Magnitude Matter? A Cross-Task Functional-Symmetry Framework

ResearchDGX agent

arXiv:2602.09229v3 Announce Type: replace Abstract: Cosine similarity normalizes both sides; dot product normalizes neither. We propose a 2x2 framework that independently controls query-side and docum

When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining

ResearchDGX agent

arXiv:2605.07756v1 Announce Type: cross Abstract: Modern deep models are often pretrained on large-scale data with missing labels using composite objectives, where the relative weights of multiple los

Why do Large Language Models Fail in Low-resource Translation? Unraveling the Token Dynamics of Large Language Models for Machine Translation

ResearchDGX agent

arXiv:2605.07533v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently demonstrated strong performance in machine translation (MT). However, most prior work focuses on improving or

You Only Stack Once (YOSO): A Motion-Filtered, Deep-Learning Framework for Detecting Faint Moving Sources

ResearchDGX agent

arXiv:2605.06913v1 Announce Type: cross Abstract: We present You Only Stack Once (YOSO), an automated pipeline designed to detect faint, slow-moving Solar System objects in wide-field astronomical sur

Zero-Shot Satellite Image Retrieval through Joint Embeddings: Application to Crisis Response

ResearchDGX agent

arXiv:2605.05405v2 Announce Type: replace Abstract: Semantic search of Earth observation archives remains challenging. Visual foundation models such as CLAY produce rich embeddings of satellite imager

10 May 2026

71% say Trump is not honest or trustworthy, and 67% say he doesn’t carefully consider important decisions — WaPo/Ipsos poll

ResearchDGX agent

A Washington Post/Ipsos poll found that 71% of respondents said Donald Trump is not honest or trustworthy, and 67% said he does not carefully consider important decisions. The poll data was shared by

Back from a little family break! Lots has happened, and I’m planning to do a deeper dive into the most interesting architectural components …

ResearchDGX agent

Back from a little family break! Lots has happened, and I’m planning to do a deeper dive into the most interesting architectural components (soon). Btw, are there any major architectures I missed belo

France is the only European country that turned nuclear generation into a structural competitive advantage. 57 reactors built between the 19…

ResearchDGX agent

France is the only European country that turned nuclear generation into a structural competitive advantage. 57 reactors built between the 1970s and 1990s produce 70% of its electricity today. Wholesal

If you cannot express your idea in the language of mathematics or code, you do not understand your idea just yet. It is only an intuition.

ResearchDGX agent

This statement by François Chollet argues that true understanding of an idea requires the ability to formalize it mathematically or through code, rather than merely having intuitive grasp. The premise

(INTENTIONALLY) LOST IN TRANSLATION: Democrats: We'd like cops to stop killing minorities. Republicans: Dems hate police. Democrats: Women s…

ResearchDGX agent

(INTENTIONALLY) LOST IN TRANSLATION: Democrats: We'd like cops to stop killing minorities. Republicans: Dems hate police. Democrats: Women should have the right to choose. Republicans: Dems want to ki

Intuition is the precursor to all important ideas.

ResearchDGX agent

François Chollet argues that intuition serves as the foundational precursor to the development of significant ideas and innovations. This perspective suggests that before formal analysis, rigorous met

防衛分野で働くSoftware Engineerのインタビュー記事を公開しました🐟 https://sakana.ai/defense-swe-interview-2026/ 内容を一部ご紹介 ・指揮統制システムや偽情報対策技術の開発など、防衛分野のSoftware Engin…

ResearchDGX agent

防衛分野で働くSoftware Engineerのインタビュー記事を公開しました🐟 https://sakana.ai/defense-swe-interview-2026/ 内容を一部ご紹介 ・指揮統制システムや偽情報対策技術の開発など、防衛分野のSoftware Engineerが担う役割 ・多様な技術領域をまたいだ開発スタイル ・ミッションクリティカルな領域での緊張感と充実感 防衛×AIの最

The Top AI Papers of the Week (May 4 - 10) - Conductor - HeavySkill - Horizon Generalization - 1,000 Synthetic Computers - Self-Improving Pr…

ResearchDGX agent

The Top AI Papers of the Week (May 4 - 10) - Conductor - HeavySkill - Horizon Generalization - 1,000 Synthetic Computers - Self-Improving Pretraining - Coordination as Architecture - Connect Four Alph

9 May 2026

It was always the case that agency was self-compounding, but AI is magnifying the effect. Low-agency AI users further lose agency, high-agen…

ResearchDGX agent

Francois Chollet argues that AI amplifies existing disparities in user agency, where individuals with high agency gain further capabilities while those with low agency become increasingly dependent, c

One theorem every ML engineer should know: The Johnson–Lindenstrauss Lemma. It states that high-dimensional data can be projected into a muc…

ResearchDGX agent

One theorem every ML engineer should know: The Johnson–Lindenstrauss Lemma. It states that high-dimensional data can be projected into a much lower-dimensional space while approximately preserving pai

Reproducing all of Schmidhuber’s papers (1990-2025) using an AI coding assistant. Cool project by @yaroslavvb! It even reproduced the “World…

ResearchDGX agent

Reproducing all of Schmidhuber’s papers (1990-2025) using an AI coding assistant. Cool project by @yaroslavvb! It even reproduced the “World Models” paper by me and @SchmidhuberAI with a toy env, with

8 May 2026

For those interested, I will be doing a live session on this topic soon: https://academy.dair.ai/events/cmovobp97000904l5h0n9a2yz Sign up if…

ResearchDGX agent

For those interested, I will be doing a live session on this topic soon: https://academy.dair.ai/events/cmovobp97000904l5h0n9a2yz Sign up if you are interested in some of the tools we are releasing so

← Previous
1…271272273274275…432
Next →