AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,872 results
27 May 2026

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.26193v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) has long been a hot research topic in data mining due to its various applications. Recent studies challenge the e

Constraint acquisition needs better benchmarks

Model ReleasesDGX agent

arXiv:2605.26279v1 Announce Type: new Abstract: Constraint Acquisition (CA) and related research on the validation and enhancement of Mathematical Programming (MP) models from domain knowledge artifac

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

AgentsDGX agent

arXiv:2509.04310v4 Announce Type: replace Abstract: Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in extit{complex}, extit

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty

SafetyDGX agent

arXiv:2605.27288v1 Announce Type: cross Abstract: Large language models (LLMs) are known to abandon their initial stance to conform to user pushback. While prior research largely attributes this behav

MATT-CTR: Unleashing a Model-Agnostic Test-Time Paradigm for CTR Prediction with Confidence-Guided Inference Paths

Model ReleasesDGX agent

arXiv:2510.08932v2 Announce Type: replace Abstract: Recently, a growing body of research has focused on either optimizing CTR model architectures to better model feature interactions or refining train

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

Model ReleasesDGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin

On the Generalization Capabilities, Design Choices and Limitations of Keypoint Imitation Learning

TutorialsDGX agent

arXiv:2605.26649v1 Announce Type: new Abstract: RGB-based imitation learning requires many demonstrations to generalize to unseen objects or scenes, motivating research into intermediate representatio

PRBench: A Standardized Probabilistic Robustness Benchmark

Model ReleasesDGX agent

arXiv:2511.01724v3 Announce Type: replace Abstract: Deep learning models are notoriously vulnerable to imperceptible perturbations. Most existing research centers on adversarial robustness (AR), which

Receipt Replay OOD: A Small Benchmark for Screen Replay Detection Under Domain Shift

Model ReleasesDGX agent

arXiv:2605.26855v1 Announce Type: new Abstract: Public datasets such as DLC-2021, SynID, and KID34K have significantly contributed to research on presentation attack detection for identity documents,

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km

Model ReleasesDGX agent

arXiv:2510.09606v2 Announce Type: replace Abstract: With the current surge in spatial reasoning explorations, researchers have made significant progress in understanding indoor scenes, but still strug

Train with autoregression & convert weights to diffusion for inference.

IndustryDGX agent

Train with autoregression & convert weights to diffusion for inference. Most researchers agree that autoregression is best when memory bandwidth is cheap and diffusion is best when FLOPS are cheap. Th

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering.…

Model ReleasesDGX agent

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering. This new research introduces AgingBench, a longitudinal rel

26 May 2026

3D-printable humanoid legs let robotics experiments run wild

IndustryDGX agent

Researchers at UC Berkeley have developed Berkeley Humanoid Lite, a low-cost, open-source robot made of 3D-printed parts , which keeps hardware costs under $5,000 with a modular design allowing easy c

AION: Next-Generation Tasks and Practical Harness for Time Series

AgentsDGX agent

arXiv:2605.25045v1 Announce Type: new Abstract: Time series research is moving beyond fixed forecasting benchmarks toward realistic tasks that combine prediction, contextual reasoning, tool use, and s

Federated Learning over Human-Body Communication for On-Body Edge Intelligence: A Survey, Taxonomy, and BODYFED-HBC Scheduling Vignette

Local AiDGX agent

arXiv:2605.24062v1 Announce Type: cross Abstract: Human-body communication (HBC) is a promising physical substrate for wearable body-area networks because it can localize communication around the body

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

Model ReleasesDGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

Human-AI Collaboration in Science at Scale: A Global Large-scale Randomized Field Experiment

ApplicationsDGX agent

arXiv:2605.24180v1 Announce Type: cross Abstract: Collaboration is the defining mode of modern science, yet its core mechanism -- feedback -- remains hard to observe, difficult to scale, and unequally

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

AgentsDGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks

SafetyDGX agent

arXiv:2605.24217v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from research environments to production deployments, evaluating their performance against strict Service Lev

Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI

SafetyDGX agent

arXiv:2605.23981v1 Announce Type: cross Abstract: Generative AI research increasingly confronts a shared problem: systems must sustain yet govern their own generative activity when uncertainty is high

Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending

Model ReleasesDGX agent

arXiv:2605.25574v1 Announce Type: cross Abstract: Concept erasure has emerged as a key research direction for ensuring safe and ethical image synthesis in Text-to-Image (T2I) models. While existing st

oh. my. god. could this word cloud diagram be … conscious?

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, questions whether a word cloud diagram could possess consciousness, likely engaging in ironic commentary on overclaimed AI capabilities or consciousn

QML-PipeGuard: Drift-Aware Behavioral Fingerprinting for Quantum Machine Learning Pipeline Integrity

SafetyDGX agent

arXiv:2605.25066v1 Announce Type: cross Abstract: Quantum machine learning (QML) is moving from research prototypes to deployed cloud services. As QML enters regulated industries, the integrity of the

STaT: Resolving Shape Distortion in Non-Stationary Time Series via Tri-Modal Synergy

SafetyDGX agent

arXiv:2605.25943v1 Announce Type: new Abstract: Recent research in time series forecasting frequently investigates the integration of textual and visual modalities with numerical models to better navi

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

SafetyDGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

25 May 2026

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

Model ReleasesDGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

Cultural Adaptation in Large Language Models for Political Discourse

Model ReleasesDGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

Design and Report Benchmarks for Knowledge Work

Model ReleasesDGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

Droneulator: A Portable UAV Simulator for Agricultural Workflows with RotorPy and Godot 4

SafetyDGX agent

arXiv:2605.23386v1 Announce Type: new Abstract: Agricultural UAV research requires simulators that integrate realistic 3D scenes, high-fidelity vehicle dynamics, and robotics middleware, while remaini

How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness

Model ReleasesDGX agent

arXiv:2605.23628v1 Announce Type: new Abstract: Multi-task benchmarks have become a central pillar of machine learning research, yet their growing influence has incentivised benchmark gaming -- strate

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

ApplicationsDGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

Physiome-ODE: A Benchmark for Irregularly Sampled Multivariate Time Series Forecasting Based on Biological ODEs

Model ReleasesDGX agent

arXiv:2502.07489v2 Announce Type: replace Abstract: State-of-the-art methods for forecasting irregularly sampled time series with missing values predominantly rely on just four datasets and a few smal

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

SafetyDGX agent

arXiv:2605.23027v1 Announce Type: new Abstract: Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where r

When Determinants Are Not Enough: Private Rare Switching

SafetyDGX agent

arXiv:2605.23131v1 Announce Type: new Abstract: In this note, I would like to share a small research moment where Codex helped me find the right way to adapt rare switching to the private setting. The

24 May 2026

even @geohotz is starting to sound like me 🤣

SafetyDGX agent

Gary Marcus humorously notes that George Hotz, an AI researcher and entrepreneur, is beginning to echo Marcus's own views or criticisms, likely regarding AI safety, limitations, or technical concerns.

not using LLM’s works for me

SafetyDGX agent

not using LLM’s works for me Prolonged AI use may make it harder to think critically and creatively, recent research suggests. But there are ways to keep the brain fit https://www.economist.com/scienc

Scientists invented a fake disease. AI told people it was real

IndustryDGX agent

Researchers from the University of Gothenburg invented a fake disease called 'bixonimania,' a fictional skin condition supposedly caused by screen time. Multiple AI chatbots including Google's Gemini,

so many experiments I want to run… 😵‍💫

AgentsDGX agent

Yohei Nakajima expresses the overwhelm of having numerous experimental ideas he wants to pursue, reflecting on the challenge of prioritization and resource constraints in AI research and development.

this is bad

TutorialsDGX agent

this is bad SHOCKING: Two researchers at Northeastern sat down with six of the chatbots that hundreds of millions of people use every day. They typed a sentence anyone in distress might type at 3 in t

23 May 2026

Discovering Entity-Conditioned Lag Heterogeneity: A Lag-Gated Neural Audit Framework for Panel Time Series

ApplicationsDGX agent

arXiv:2605.21542v1 Announce Type: new Abstract: Country-level temporal panels are widely used in empirical analysis. Researchers often need to audit how different entities respond to historical signal

During his second term, Trump will have cut 2 of the US most powerful innovation and wealth creation engines: 1. skilled legal immigration 2…

SafetyDGX agent

During his second term, Trump will have cut 2 of the US most powerful innovation and wealth creation engines: 1. skilled legal immigration 2. (non-defense) research budgets We won't see the effect of

If you can learn one thing that's genuinely novel to you, you can learn anything.

TutorialsDGX agent

This statement from AI researcher François Chollet suggests that the ability to learn something genuinely novel demonstrates a fundamental learning capacity that generalizes across all domains. The cl

22 May 2026

Does Slightly Mean Somewhat? Measuring Vague Intensity Words in LLM Numeric Actions

Model ReleasesDGX agent

arXiv:2605.21827v1 Announce Type: new Abstract: Do language models preserve the ordinal meaning of intensity words when those words must produce numeric actions? I study a researcher-constructed scale

nope you are. because openai will sell your most private data to the government.

SafetyDGX agent

This post appears to express concerns about OpenAI's data privacy practices and potential government data sharing, though the claim lacks specific evidence or context. Gary Marcus, an AI researcher an

21 May 2026

AI-based Prediction of Independent Construction Safety Outcomes from Universal Attributes

SafetyDGX agent

arXiv:1908.05972v3 Announce Type: replace Abstract: This paper significantly improves on, and finishes to validate, an approach proposed in previous research in which safety outcomes were predicted fr

Capability neq Interpretability: Human Interpretability of Vision Foundation Models

Model ReleasesDGX agent

arXiv:2605.20337v1 Announce Type: new Abstract: How interpretable are the features of leading vision models? The question is increasingly pressing as these models move from research benchmarks into hi

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

SafetyDGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI

HardwareDGX agent

At NVIDIA GTC Taipei at COMPUTEX, the world’s developers, researchers and industry leaders are converging to dive into the latest breakthroughs shaping every industry, covering topics spanning AI fact

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2605.21261v1 Announce Type: new Abstract: Training-free zero-shot composed image retrieval models are recently gaining increasing research interest due to their generalizability and flexibility

20 May 2026

A Logistic Regression Model to Predict Malaria Severity in Children

AgentsDGX agent

arXiv:2605.18900v1 Announce Type: cross Abstract: One of the main causes of death around the globe is malaria. Researchers have sought to develop predictive models for malaria outbreaks based on meteo

Agent Security is a Systems Problem

AgentsDGX agent

arXiv:2605.18991v1 Announce Type: cross Abstract: We take the position that agent security must be approached as a systems problem: the AI model powering the agent must be treated as an untrusted comp

AgentNLQ: A General-Purpose Agent for Natural Language to SQL

Model ReleasesDGX agent

arXiv:2605.19010v1 Announce Type: new Abstract: Natural language to SQL (NL2SQL) conversion is an important problem for researchers and enterprises due to the ubiquitous importance of relational datab

Cardiac fat segmentation using computed tomography and an image-to-image conditional generative adversarial neural network

AgentsDGX agent

arXiv:2605.20064v1 Announce Type: new Abstract: In recent years, research has highlighted the association between increased adipose tissue surrounding the human heart and elevated susceptibility to ca

Causal Evidence that Language Models use Confidence to Drive Behavior

AgentsDGX agent

arXiv:2603.22161v2 Announce Type: replace Abstract: Metacognition -- assessing the quality of one's own cognitive performance -- guides adaptive behavior across species. Substantial research demonstra

Context compression isn't new in RAG. Our contribution is making it query-aware, citation-preserving, and fast enough for orchestration. Rea…

ToolsDGX agent

Context compression isn't new in RAG. Our contribution is making it query-aware, citation-preserving, and fast enough for orchestration. Read the full research blog: https://research.perplexity.ai/art

Data-Efficient Self-Supervised Algorithms for Fine-Grained Birdsong Analysis

ApplicationsDGX agent

arXiv:2511.12158v3 Announce Type: replace Abstract: Research in bioacoustics, neuroscience, and linguistics often uses birdsong as a proxy to acquire knowledge across diverse areas. This requires audi

Distributional AGI Safety

SafetyDGX agent

arXiv:2512.16856v2 Announce Type: replace Abstract: AI safety and alignment research has predominantly been focused on methods for safeguarding individual AI systems, resting on the assumption of an e

For centuries, the scientific method has been our best tool for progress. But today, there’s so much data out there that it’s impossible for…

Model ReleasesDGX agent

For centuries, the scientific method has been our best tool for progress. But today, there’s so much data out there that it’s impossible for any one researcher to connect all the dots. We want to fix

GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation

Model ReleasesDGX agent

arXiv:2605.19890v1 Announce Type: new Abstract: Test-time adaptation (TTA) enables a pre-trained model to adapt online to an unlabeled test stream under distribution shift. While most TTA research foc

← Previous
1…355356357358359…432
Next →