AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,880 results
Model Releases

Snippet-Driven Supply Chain Discovery with LLMs: Scaling Visibility in China

DGX agent

arXiv:2605.27845v1 Announce Type: cross Abstract: Financial and economic research often relies on structured supply-chain disclosures and commercial databases. In China, supplier--customer disclosure

model-releasesarxiv-cs-ai
28 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Trust Me, I'm an Expert: Decoding and Steering Authority Bias in Large Language Models

DGX agent

arXiv:2601.13433v3 Announce Type: replace Abstract: Prior research demonstrates that performance of language models on reasoning tasks can be influenced by suggestions, hints and endorsements. However

safetyarxiv-cs-cl
28 May 2026
Model Releases

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequo…

DGX agent

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequoia. This investment will help us advance our research and expa

model-releasessonya-huang--x
28 May 2026
Model Releases

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

DGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection

DGX agent

arXiv:2605.26193v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) has long been a hot research topic in data mining due to its various applications. Recent studies challenge the e

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Constraint acquisition needs better benchmarks

DGX agent

arXiv:2605.26279v1 Announce Type: new Abstract: Constraint Acquisition (CA) and related research on the validation and enhancement of Mathematical Programming (MP) models from domain knowledge artifac

model-releasesarxiv-cs-ai
27 May 2026
Agents

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

DGX agent

arXiv:2509.04310v4 Announce Type: replace Abstract: Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in extit{complex}, extit

agentsarxiv-cs-ai
27 May 2026
Safety

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty

DGX agent

arXiv:2605.27288v1 Announce Type: cross Abstract: Large language models (LLMs) are known to abandon their initial stance to conform to user pushback. While prior research largely attributes this behav

safetyarxiv-cs-ai
27 May 2026
Model Releases

MATT-CTR: Unleashing a Model-Agnostic Test-Time Paradigm for CTR Prediction with Confidence-Guided Inference Paths

DGX agent

arXiv:2510.08932v2 Announce Type: replace Abstract: Recently, a growing body of research has focused on either optimizing CTR model architectures to better model feature interactions or refining train

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Modeling Dynamic Mixtures of Time-Delay Systems from Streaming Time Series

DGX agent

arXiv:2605.26191v1 Announce Type: cross Abstract: This research addresses the problem of adaptive modeling in time-series data streams with clear input-output relationships. This problem is challengin

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

On the Generalization Capabilities, Design Choices and Limitations of Keypoint Imitation Learning

DGX agent

arXiv:2605.26649v1 Announce Type: new Abstract: RGB-based imitation learning requires many demonstrations to generalize to unseen objects or scenes, motivating research into intermediate representatio

tutorialsarxiv-cs-ro
27 May 2026
Model Releases

PRBench: A Standardized Probabilistic Robustness Benchmark

DGX agent

arXiv:2511.01724v3 Announce Type: replace Abstract: Deep learning models are notoriously vulnerable to imperceptible perturbations. Most existing research centers on adversarial robustness (AR), which

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Receipt Replay OOD: A Small Benchmark for Screen Replay Detection Under Domain Shift

DGX agent

arXiv:2605.26855v1 Announce Type: new Abstract: Public datasets such as DLC-2021, SynID, and KID34K have significantly contributed to research on presentation attack detection for identity documents,

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km

DGX agent

arXiv:2510.09606v2 Announce Type: replace Abstract: With the current surge in spatial reasoning explorations, researchers have made significant progress in understanding indoor scenes, but still strug

model-releasesarxiv-cs-cv
27 May 2026
Industry

Train with autoregression & convert weights to diffusion for inference.

DGX agent

Train with autoregression & convert weights to diffusion for inference. Most researchers agree that autoregression is best when memory bandwidth is cheap and diffusion is best when FLOPS are cheap. Th

industryemad-mostaque--x
27 May 2026
Model Releases

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering.…

DGX agent

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering. This new research introduces AgingBench, a longitudinal rel

model-releasesdair-ai--x
27 May 2026
Industry

3D-printable humanoid legs let robotics experiments run wild

DGX agent

Researchers at UC Berkeley have developed Berkeley Humanoid Lite, a low-cost, open-source robot made of 3D-printed parts , which keeps hardware costs under $5,000 with a modular design allowing easy c

industryars-technica
26 May 2026
Agents

AION: Next-Generation Tasks and Practical Harness for Time Series

DGX agent

arXiv:2605.25045v1 Announce Type: new Abstract: Time series research is moving beyond fixed forecasting benchmarks toward realistic tasks that combine prediction, contextual reasoning, tool use, and s

agentsarxiv-cs-ai
26 May 2026
Local Ai

Federated Learning over Human-Body Communication for On-Body Edge Intelligence: A Survey, Taxonomy, and BODYFED-HBC Scheduling Vignette

DGX agent

arXiv:2605.24062v1 Announce Type: cross Abstract: Human-body communication (HBC) is a promising physical substrate for wearable body-area networks because it can localize communication around the body

local-aiarxiv-cs-ai
26 May 2026
Model Releases

From Index to Equity: Pre-Training Transformers for Stock Return Prediction

DGX agent

arXiv:2605.23962v1 Announce Type: cross Abstract: This research aims to leverage machine learning to improve stock price prediction and support informed investment decisions related to buying, selling

model-releasesarxiv-cs-lg
26 May 2026
Applications

Human-AI Collaboration in Science at Scale: A Global Large-scale Randomized Field Experiment

DGX agent

arXiv:2605.24180v1 Announce Type: cross Abstract: Collaboration is the defining mode of modern science, yet its core mechanism -- feedback -- remains hard to observe, difficult to scale, and unequally

applicationsarxiv-cs-ai
26 May 2026
Agents

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

DGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

agentsarxiv-cs-ai
26 May 2026
Safety

Identifying and Mitigating Systemic Measurement Bias in Production LLM Inference Benchmarks

DGX agent

arXiv:2605.24217v1 Announce Type: new Abstract: As Large Language Models (LLMs) transition from research environments to production deployments, evaluating their performance against strict Service Lev

safetyarxiv-cs-ai
26 May 2026
Safety

Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governance in Generative AI

DGX agent

arXiv:2605.23981v1 Announce Type: cross Abstract: Generative AI research increasingly confronts a shared problem: systems must sustain yet govern their own generative activity when uncertainty is high

safetyarxiv-cs-ai
26 May 2026
Model Releases

Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending

DGX agent

arXiv:2605.25574v1 Announce Type: cross Abstract: Concept erasure has emerged as a key research direction for ensuring safe and ethical image synthesis in Text-to-Image (T2I) models. While existing st

model-releasesarxiv-cs-ai
26 May 2026
Safety

oh. my. god. could this word cloud diagram be … conscious?

DGX agent

Gary Marcus, a prominent AI researcher and critic, questions whether a word cloud diagram could possess consciousness, likely engaging in ironic commentary on overclaimed AI capabilities or consciousn

safetygary-marcus--x
26 May 2026
Safety

QML-PipeGuard: Drift-Aware Behavioral Fingerprinting for Quantum Machine Learning Pipeline Integrity

DGX agent

arXiv:2605.25066v1 Announce Type: cross Abstract: Quantum machine learning (QML) is moving from research prototypes to deployed cloud services. As QML enters regulated industries, the integrity of the

safetyarxiv-cs-lg
26 May 2026
Safety

STaT: Resolving Shape Distortion in Non-Stationary Time Series via Tri-Modal Synergy

DGX agent

arXiv:2605.25943v1 Announce Type: new Abstract: Recent research in time series forecasting frequently investigates the integration of textual and visual modalities with numerical models to better navi

safetyarxiv-cs-lg
26 May 2026
Safety

TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation

DGX agent

arXiv:2605.25547v1 Announce Type: new Abstract: Existing embodied control research demonstrates remarkable performance improvements by scaling training data and model size. We instead explore inferenc

safetyarxiv-cs-ro
26 May 2026
Model Releases

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

DGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Cultural Adaptation in Large Language Models for Political Discourse

DGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Design and Report Benchmarks for Knowledge Work

DGX agent

arXiv:2605.23262v1 Announce Type: new Abstract: The development of LLM agents has led to a growing body of work on knowledge-work AI, including coding, research, and healthcare. However, current knowl

model-releasesarxiv-cs-ai
25 May 2026
Safety

Droneulator: A Portable UAV Simulator for Agricultural Workflows with RotorPy and Godot 4

DGX agent

arXiv:2605.23386v1 Announce Type: new Abstract: Agricultural UAV research requires simulators that integrate realistic 3D scenes, high-fidelity vehicle dynamics, and robotics middleware, while remaini

safetyarxiv-cs-ro
25 May 2026
Model Releases

How Hard is it to Rig a Benchmark? A Social Choice Analysis of Leaderboard Robustness

DGX agent

arXiv:2605.23628v1 Announce Type: new Abstract: Multi-task benchmarks have become a central pillar of machine learning research, yet their growing influence has incentivised benchmark gaming -- strate

model-releasesarxiv-cs-lg
25 May 2026
Applications

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

DGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

applicationsarxiv-cs-cl
25 May 2026
Model Releases

Physiome-ODE: A Benchmark for Irregularly Sampled Multivariate Time Series Forecasting Based on Biological ODEs

DGX agent

arXiv:2502.07489v2 Announce Type: replace Abstract: State-of-the-art methods for forecasting irregularly sampled time series with missing values predominantly rely on just four datasets and a few smal

model-releasesarxiv-cs-lg
25 May 2026
Safety

PIMbot: A Self-Adaptive Attack Framework for Adversarial Manipulation of Multi-Robot Reinforcement Learning

DGX agent

arXiv:2605.23027v1 Announce Type: new Abstract: Recent research has demonstrated the potential of reinforcement learning in effective multi-robot collaboration, particularly in social dilemmas where r

safetyarxiv-cs-ro
25 May 2026
Safety

When Determinants Are Not Enough: Private Rare Switching

DGX agent

arXiv:2605.23131v1 Announce Type: new Abstract: In this note, I would like to share a small research moment where Codex helped me find the right way to adapt rare switching to the private setting. The

safetyarxiv-cs-lg
25 May 2026
Safety

even @geohotz is starting to sound like me 🤣

DGX agent

Gary Marcus humorously notes that George Hotz, an AI researcher and entrepreneur, is beginning to echo Marcus's own views or criticisms, likely regarding AI safety, limitations, or technical concerns.

safetygary-marcus--x
24 May 2026
Safety

not using LLM’s works for me

DGX agent

not using LLM’s works for me Prolonged AI use may make it harder to think critically and creatively, recent research suggests. But there are ways to keep the brain fit https://www.economist.com/scienc

safetygary-marcus--x
24 May 2026
Industry

Scientists invented a fake disease. AI told people it was real

DGX agent

Researchers from the University of Gothenburg invented a fake disease called 'bixonimania,' a fictional skin condition supposedly caused by screen time. Multiple AI chatbots including Google's Gemini,

industryr-chatgpt
24 May 2026
Agents

so many experiments I want to run… 😵‍💫

DGX agent

Yohei Nakajima expresses the overwhelm of having numerous experimental ideas he wants to pursue, reflecting on the challenge of prioritization and resource constraints in AI research and development.

agentsyohei-nakajima--x
24 May 2026
Tutorials

this is bad

DGX agent

this is bad SHOCKING: Two researchers at Northeastern sat down with six of the chatbots that hundreds of millions of people use every day. They typed a sentence anyone in distress might type at 3 in t

tutorialsgary-marcus--x
24 May 2026
Applications

Discovering Entity-Conditioned Lag Heterogeneity: A Lag-Gated Neural Audit Framework for Panel Time Series

DGX agent

arXiv:2605.21542v1 Announce Type: new Abstract: Country-level temporal panels are widely used in empirical analysis. Researchers often need to audit how different entities respond to historical signal

applicationsarxiv-cs-lg
23 May 2026
Safety

During his second term, Trump will have cut 2 of the US most powerful innovation and wealth creation engines: 1. skilled legal immigration 2…

DGX agent

During his second term, Trump will have cut 2 of the US most powerful innovation and wealth creation engines: 1. skilled legal immigration 2. (non-defense) research budgets We won't see the effect of

safetyyann-lecun--x
23 May 2026
Tutorials

If you can learn one thing that's genuinely novel to you, you can learn anything.

DGX agent

This statement from AI researcher François Chollet suggests that the ability to learn something genuinely novel demonstrates a fundamental learning capacity that generalizes across all domains. The cl

tutorialsfrancois-chollet--x
23 May 2026
Model Releases

Does Slightly Mean Somewhat? Measuring Vague Intensity Words in LLM Numeric Actions

DGX agent

arXiv:2605.21827v1 Announce Type: new Abstract: Do language models preserve the ordinal meaning of intensity words when those words must produce numeric actions? I study a researcher-constructed scale

model-releasesarxiv-cs-cl
22 May 2026
Safety

nope you are. because openai will sell your most private data to the government.

DGX agent

This post appears to express concerns about OpenAI's data privacy practices and potential government data sharing, though the claim lacks specific evidence or context. Gary Marcus, an AI researcher an

safetygary-marcus--x
22 May 2026
← Previous
1…444445446447448…540
Next →