AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,872 results
20 May 2026

JAXenstein: Accelerated Benchmarking for First-Person Environments

Model ReleasesDGX agent

arXiv:2605.19926v1 Announce Type: new Abstract: The progression of reinforcement learning algorithms have been driven by challenging benchmarks. The rate in which a researcher can iterate on a problem

Operationalizing Document AI: A Microservice Architecture for OCR and LLM Pipelines in Production

Model ReleasesDGX agent

arXiv:2605.18818v1 Announce Type: new Abstract: Academic research tends to focus on new models for document understanding creating a wide gap in the literature between model definition and running mod

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

Model ReleasesDGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

This result points to something larger: AI systems are becoming capable of holding together long, difficult chains of reasoning, connecting …

Model ReleasesDGX agent

This result points to something larger: AI systems are becoming capable of holding together long, difficult chains of reasoning, connecting ideas across distant fields, and surfacing paths researchers

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

Model ReleasesDGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

We are also launching Science Skills, a specialized bundle that integrates insights from 30+ major life science models and databases with ag…

AgentsDGX agent

We are also launching Science Skills, a specialized bundle that integrates insights from 30+ major life science models and databases with agentic platforms like @Antigravity to allow researchers to pe

19 May 2026

Actionable World Representation

SafetyDGX agent

arXiv:2605.18743v1 Announce Type: new Abstract: Inspired by the emergent behaviors in large language models that generalized human intelligence, the research community is pursuing similar emergent cap

[AINews] How to land a job at a frontier lab (on Pretraining)

TutorialsDGX agent

This article from Latent Space provides guidance on securing employment at cutting-edge AI research laboratories, with a focus on roles related to pretraining large language models. It likely covers p

CheckSupport: A Local LLM-Powered Tool for Automated Manuscript Submission Checklist Selection and Completion

Local AiDGX agent

arXiv:2605.16377v1 Announce Type: cross Abstract: Transparent and standardized reporting is essential for reproducible scientific research, yet adherence to reporting guidelines remains inconsistent b

Demis Hassabis said this might be the ‘foothills of the singularity.’ What?

IndustryDGX agent

Welcome to a 'profound moment for humanity,' according to Google DeepMind CEO Demis Hassabis, who closed out Google I/O's keynote presentation on Tuesday, saying: Google's cutting-edge research and pr

Democratizing Large-Scale Re-Optimization with LLM-Guided Model Patches

AgentsDGX agent

arXiv:2605.18692v1 Announce Type: new Abstract: Optimization models developed by operations research (OR) experts are often deployed as decision-support systems in industrial settings. However, real-w

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap

SafetyDGX agent

arXiv:2508.04149v2 Announce Type: replace-cross Abstract: Aligning large language models (LLMs) with human preferences is a critical challenge in AI research. While methods like Reinforcement Learning

DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2405.19189v3 Announce Type: replace Abstract: With the great success of diffusion models (DMs) in generating realistic synthetic vision data, many researchers have investigated their potential i

EgoKit: Towards Unified Low-Cost Egocentric Data Collection with Heterogeneous Devices

Local AiDGX agent

arXiv:2605.16797v1 Announce Type: new Abstract: Egocentric video is increasingly used as a data source for robot learning, activity understanding, and embodied AI research, but collecting it at scale

Forget Many, Forget Right: Scalable and Precise Concept Unlearning in Diffusion Models

SafetyDGX agent

arXiv:2601.06162v4 Announce Type: replace-cross Abstract: Text-to-image diffusion models have achieved remarkable progress, yet their use raises copyright and misuse concerns, prompting research into

Harnessing AI for Inverse Partial Differential Equation Problems: Past, Present, and Prospects

ApplicationsDGX agent

arXiv:2605.16966v1 Announce Type: new Abstract: Solving inverse partial differential equation (PDE) problems is a fundamental topic in scientific research due to its broad significance across a wide r

Herding CATs: ALARA for Agent Harness Engineering in Portable Composable Multi-Agent Teams

Local AiDGX agent

arXiv:2603.20380v2 Announce Type: replace-cross Abstract: Industry practitioners and academic researchers regularly use multi-agent systems to accelerate their work, but the applications through which

LLMs for automatic annotation of Mandarin narrative transcripts

Local AiDGX agent

arXiv:2605.17205v1 Announce Type: new Abstract: Linguistic annotation of transcribed speech is essential for research in language acquisition, language disorders, and sociolinguistics, yet remains lab

ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics

Model ReleasesDGX agent

arXiv:2605.18617v1 Announce Type: cross Abstract: Most existing vision-language manipulation research targets rigid robotic arms, whose fixed morphology limits adaptability in cluttered or confined sp

Multi-Object Tracking Consistently Improves Wildlife Inference

ApplicationsDGX agent

arXiv:2605.16672v1 Announce Type: cross Abstract: Camera traps have become a common tool for wildlife monitoring efforts in ecological research and biodiversity conservation. Wildlife classification m

OmniCode: A Benchmark for Evaluating Software Engineering Agents

Model ReleasesDGX agent

arXiv:2602.02262v3 Announce Type: replace-cross Abstract: LLM-powered coding agents are redefining how real-world software is developed. To drive the research towards better coding agents, we require

PySIFT: GPU-Resident Deterministic SIFT for Deep Learning Vision Pipelines

Local AiDGX agent

arXiv:2605.17869v1 Announce Type: new Abstract: A widespread assumption in local feature research holds that classical handcrafted descriptors are accuracy-limited relics best replaced by learned alte

Responsible Federated LLMs via Safety Filtering and Constitutional AI

SafetyDGX agent

arXiv:2502.16691v2 Announce Type: replace Abstract: Recent research has increasingly focused on training large language models (LLMs) using federated learning, known as FedLLM. However, responsible AI

SCOUT: Cyclic Causal Discovery Under Soft Interventions with Unknown Targets

ApplicationsDGX agent

arXiv:2605.16620v1 Announce Type: new Abstract: Learning causal relationships between variables from data is a fundamental research area with many applications across disciplines. Most existing causal

State-of-the-Art Claims Require State-of-the-Art Evidence

Model ReleasesDGX agent

arXiv:2605.17273v1 Announce Type: cross Abstract: State-of-the-Art (SOTA) claims pervade Artificial Intelligence (AI) and Machine Learning (ML) research. These claims rest on benchmark evaluations, wh

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

HardwareDGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

The MixCount Dataset: Bridging the Data Gap for Open-Vocabulary Object Counting

Model ReleasesDGX agent

arXiv:2605.18063v1 Announce Type: new Abstract: Object counting is a foundational vision task with over a decade of dedicated research, yet state-of-the-art models still fail systematically in the mix

Threats to Arabic Handwriting Recognition: Investigating Black-Box Adversarial Attacks on embedded ConvNet models

Model ReleasesDGX agent

arXiv:2605.18058v1 Announce Type: new Abstract: Arabic handwriting recognition (AHR) has made significant progress with deep learning models. AHR research has largely focused on performance, with secu

Tongyi DeepResearch Technical Report

AgentsDGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens

Model ReleasesDGX agent

arXiv:2605.16638v1 Announce Type: new Abstract: Recent research has demonstrated that Universal Multimodal Embedding (UME) benefits significantly from Chain-of-Thought (CoT) reasoning. In this paradig

Vision Transformer-Conditioned UNet for Domain-Adaptive Semantic Segmentation

SafetyDGX agent

arXiv:2605.16393v1 Announce Type: cross Abstract: Semantic segmentation is essential for analysing anatomical features in biomedical research, yet a performance gap remains for Vision Transformers (Vi

18 May 2026

Benchmark of Benchmarks: Unpacking Influence and Code Repository Quality in LLM Safety Benchmarks

Model ReleasesDGX agent

arXiv:2603.04459v3 Announce Type: replace-cross Abstract: The rapid expansion of research in LLM safety presents challenges in tracking advancements, making benchmarks important evaluation infrastruct

Forcepoint details TeamPCP supply chain attack that turned LiteLLM into a credential stealer

IndustryDGX agent

A new report out today from cybersecurity company Forcepoint LLC’s X-Labs research team details a supply chain attack that compromised LiteLLM, a widely used open-source Python library that serves as

Health-Conditioned Vision-Language-Action Models for Malfunction-Aware Robot Control

TutorialsDGX agent

arXiv:2605.16056v1 Announce Type: new Abstract: Research on Vision Language Action (VLA) models has been increasing rapidly in recent years. Although some of them focus on detecting, preventing, and r

Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in …

AgentsDGX agent

Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in real-time. Try our playable research preview today, with Ago

Layer Equivalence Is Not a Property of Layers Alone: How You Test Redundancy Changes What You Find

Model ReleasesDGX agent

arXiv:2605.16234v1 Announce Type: cross Abstract: When researchers ask whether two transformer layers are 'equivalent' for compression, they often conflate distinct tests. Replacement asks whether one

STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System

SafetyDGX agent

arXiv:2605.16137v1 Announce Type: new Abstract: Generating simulation-ready tabletop scenes from task instructions is an intriguing and promising research direction in the field of Embodied AI. Howeve

17 May 2026

A lot of the discussion of the good and bad of phones on society is really a discussion about the end of boredom. Boredom makes people seek …

ApplicationsDGX agent

A lot of the discussion of the good and bad of phones on society is really a discussion about the end of boredom. Boredom makes people seek out things to do. That has both good & bad effects: research

Probably one of the best culture references wrt to AI progress is the 2024 AI Film Festival poster

IndustryDGX agent

The 2024 AI Film Festival poster serves as a significant cultural reference point for documenting and reflecting on AI progress, according to AI researcher Cristobal Valenzuela. The poster likely show

16 May 2026

Heading to hashtag#MLSys2026? Come unwind with the Together AI team at Inference After Dark. Drinks, bites, shuffleboard, and a room full of…

ToolsDGX agent

Heading to hashtag#MLSys2026? Come unwind with the Together AI team at Inference After Dark. Drinks, bites, shuffleboard, and a room full of researchers and AI-native builders. 🟠 Tuesday, May 19 🟠 7:3

tesla robotaxis going about as well you might expect.

SafetyDGX agent

Gary Marcus, a prominent AI researcher and critic, comments on Tesla's robotaxi development progress, likely offering skeptical or cautionary observations about the practical challenges and timeline o

15 May 2026

AI-assisted cultural heritage dissemination: Comparing NMT and glossary-augmented LLM translation in rock art documents

Model ReleasesDGX agent

arXiv:2605.14679v1 Announce Type: cross Abstract: Cultural heritage institutions increasingly disseminate research and interpretive materials globally, but multilingual dissemination is constrained by

ARES-LSHADE: Autoresearch-Enhanced LSHADE with Memetic Polish for the GNBG Benchmark

Model ReleasesDGX agent

arXiv:2605.13877v1 Announce Type: cross Abstract: We present ARES-LSHADE, a memetic differential-evolution variant submitted to the GECCO 2026 competition on LLM-designed evolutionary algorithms for t

CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation

Model ReleasesDGX agent

arXiv:2602.20571v2 Announce Type: replace Abstract: Many benchmarks for automated causal inference evaluate a system's performance based on a single numerical output, such as an Average Treatment Effe

Identifying Culprits Through Deep Deterministic Policy Gradient Deep Learning Investigation

SafetyDGX agent

arXiv:2605.14774v1 Announce Type: new Abstract: In the world of AI and advanced technologies investigation aspects identification of a crime or criminal plays a major problem. In this research we focu

𝕏 just open-sourced its For You algorithm. This is one of the biggest transparency moves from any social platform. Here’s the Grok breakdow…

IndustryDGX agent

𝕏 (formerly Twitter) open-sourced its 'For You' recommendation algorithm, marking a significant transparency initiative by the social platform. This move allows developers and researchers to examine h

MathAtlas: A Benchmark for Autoformalization in the Wild

Model ReleasesDGX agent

arXiv:2605.14061v1 Announce Type: new Abstract: Current autoformalization benchmarks are largely focused on olympiad or undergraduate mathematics, while graduate and research-level mathematics remains

Mixed Integer Goal Programming for Personalized Meal Optimization with User-Defined Serving Granularity

Model ReleasesDGX agent

arXiv:2605.13849v1 Announce Type: new Abstract: Determining what to eat to satisfy nutritional requirements is one of the oldest optimization problems in operations research, yet existing formulations

OPT-Engine: Benchmarking the Limits of LLMs in Optimization Modeling via Complexity Scaling

Model ReleasesDGX agent

arXiv:2601.19924v2 Announce Type: replace-cross Abstract: We investigate the capabilities and scalability of Large Language Models (LLMs) in optimization modeling, a domain requiring structured reason

Routine vaccines may cut dementia risk—experts have startling hypothesis on how

IndustryDGX agent

Multiple observational studies have found that routine adult vaccines are associated with a reduced dementia risk, with some showing risk reductions of 25% to 40%. Researchers hypothesize that certain

Run @NousResearch's Hermes Agent fully locally on DGX Spark. 🚀 Our newest playbook shows you how to get set up via @Ollama step by step. 👇

Local AiDGX agent

This playbook provides step-by-step instructions for running Nous Research's Hermes Agent locally on NVIDIA DGX Spark using Ollama, enabling users to deploy an open-source AI agent entirely on local h

Solar power production undercut by coal pollution

ApplicationsDGX agent

New research reveals that pollution from coal-fired power plants is significantly reducing the energy output of solar installations, particularly where coal and solar capacity expand side by side. Aer

Teaching and Evaluating LLMs to Reason About Polymer Design Related Tasks

Model ReleasesDGX agent

arXiv:2601.16312v2 Announce Type: replace-cross Abstract: Research in AI4Science has shown promise in many science applications, including polymer design. However, current LLMs are ineffective in this

TILBench: A Systematic Benchmark for Tabular Imbalanced Learning Across Data Regimes

Model ReleasesDGX agent

arXiv:2605.14915v1 Announce Type: new Abstract: Imbalanced learning remains a fundamental challenge in tabular data applications. Despite decades of research and numerous proposed algorithms, a system

XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations

SafetyDGX agent

arXiv:2511.02776v2 Announce Type: replace Abstract: Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. Howe

14 May 2026

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

SafetyDGX agent

arXiv:2510.08992v3 Announce Type: replace Abstract: While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure th

Establishing AI and data sovereignty in the age of autonomous systems

AgentsDGX agent

When generative AI first moved from research labs into real-world business applications, enterprises made a tacit bargain: “Capability now, control later.” Feed your proprietary data into third-party

Is Video Anomaly Detection Misframed? Evidence from LLM-Based and Multi-Scene Models

SafetyDGX agent

arXiv:2605.12725v1 Announce Type: new Abstract: Recent video anomaly detection research has expanded rapidly with an emphasis on general models of normality intended to work across many different scen

Making humans responsible for their AI use seems like an incredibly reasonable way to address problems & opportunities in the use of AI for …

AgentsDGX agent

Making humans responsible for their AI use seems like an incredibly reasonable way to address problems & opportunities in the use of AI for academic research, at least in the short term (autonomous sc

← Previous
1…356357358359360…432
Next →