AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
25 May 2026

How Far Will They Go? Red-Teaming Online Influence with Large Language Models

Local AiDGX agent

arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam

ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization

Model ReleasesDGX agent

arXiv:2605.22885v1 Announce Type: new Abstract: Formal mathematics libraries are rapidly expanding, creating a growing need to refactor verified proofs for maintainability and to improve training data

Less Effort, Shorter Proofs: Reinforcement Learning for Security Protocol Analysis in Tamarin

TutorialsDGX agent

arXiv:2605.23643v1 Announce Type: cross Abstract: Tools like Tamarin and ProVerif have achieved notable success in analyzing and verifying complex real-world protocols such as EMV, 5G, and WPA2, even

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NLG Evaluation: Past, Present, Future

SafetyDGX agent

arXiv:2605.23715v1 Announce Type: new Abstract: Natural Language Generation (NLG) evaluation has changed dramatically since 1990, and will continue to evolve in the future. In 1990, when NLG had close

Online Partitioned Local Depth for semi-supervised applications

Model ReleasesDGX agent

arXiv:2512.15436v2 Announce Type: replace-cross Abstract: We introduce an extension of the partitioned local depth (PaLD) algorithm that is adapted to online applications such as semi-supervised predi

Open Multimodal Datasets and Open-Source Software for Data-Driven Modeling of Multiphase Transport and Thermal Systems

Model ReleasesDGX agent

arXiv:2605.23037v1 Announce Type: new Abstract: Data-driven modeling is becoming central to multiphase transport, electronics cooling, acoustic diagnostics, and thermal-fluid digital twins, but progre

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control

SafetyDGX agent

arXiv:2605.22894v1 Announce Type: cross Abstract: Controlling physics-based humanoids from natural-language instructions is a critical step toward general-purpose embodied agents. However, existing me

TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2602.01665v2 Announce Type: replace-cross Abstract: The design of environments plays a critical role in shaping the development and evaluation of cooperative multi-agent reinforcement learning (

The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems

ApplicationsDGX agent

arXiv:2605.23024v1 Announce Type: new Abstract: Large language models now write software, draft legal documents, and produce clinical notes, yet fundamental limits, from Turing and Arrow to the No Fre

Topological Signal Processing: An Application-Oriented Tutorial

TutorialsDGX agent

arXiv:2605.22853v1 Announce Type: cross Abstract: Many modern datasets are large and carry complex structural relationships. Graph-based methods have traditionally been used to represent networked dat

Understanding and Improving Noisy Embedding Techniques in Instruction Finetuning

Model ReleasesDGX agent

arXiv:2605.23171v1 Announce Type: cross Abstract: Recent advancements in instructional fine-tuning have injected noise into embeddings, with NEFTune (Jain et al., 2024) setting benchmarks using unifor

VideoOdyssey: A Benchmark for Ultra-Long-Context and Omni-Modal Video Understanding

Model ReleasesDGX agent

arXiv:2605.22907v1 Announce Type: new Abstract: Real-world long video understanding requires models to perform continuous tracking, information integration and memory retention over massive temporal s

24 May 2026

Moment, which develops AI tools for automating fixed-income and equities trading tech, raised a $78M Series C led by Index Ventures, with a16z participating (Paige Smith/Bloomberg)

IndustryDGX agent

Paige Smith / Bloomberg: Moment, which develops AI tools for automating fixed-income and equities trading tech, raised a $78M Series C led by Index Ventures, with a16z participating — Moment, the fina

The only thing growing faster than the artificial-intelligence industry may be Americans’ negative feelings about it. https://on.wsj.com/3PT…

SafetyDGX agent

A Wall Street Journal report highlights the paradox that while the artificial intelligence industry experiences rapid growth, American public sentiment toward AI is increasingly negative. The article,

23 May 2026

DecepChain: Inducing Deceptive Reasoning in Large Language Models

SafetyDGX agent

arXiv:2510.00319v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been demonstrating strong reasoning capability with their chain-of-thoughts (CoT), which are routinely used by hum

FD-Bench: A Modular and Fair Benchmark for Data-driven Fluid Simulation

Model ReleasesDGX agent

arXiv:2505.20349v2 Announce Type: replace-cross Abstract: Data-driven modeling of fluid dynamics has advanced rapidly with neural PDE solvers, yet a fair and strong benchmark remains fragmented due to

HealthMamba: An Uncertainty-aware Spatiotemporal Graph State Space Model for Effective and Reliable Healthcare Facility Visit Prediction

SafetyDGX agent

arXiv:2602.05286v3 Announce Type: replace Abstract: Healthcare facility visit prediction is essential for optimizing healthcare resource allocation and informing public health policy. Despite advanced

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

HardwareDGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

Lumberjack: Better Differentially Private Random Forests through Heavy Hitter Detection in Trees

Model ReleasesDGX agent

arXiv:2605.22756v1 Announce Type: new Abstract: Random forests are widely used in fields involving sensitive tabular data, but existing approaches to enforcing differential privacy (DP) typically degr

Position: The Time for Sampling Is Now! Charting a New Course for Bayesian Deep Learning

TutorialsDGX agent

arXiv:2605.21765v1 Announce Type: new Abstract: The practical adoption of sampling-based inference (SAI) in Bayesian neural networks (BNNs) remains limited, partly due to persistent misconceptions abo

six years ago world (cognitive) models were the centerpiece of my essay The Next Decade in AI. their time is finally coming.

SafetyDGX agent

six years ago world (cognitive) models were the centerpiece of my essay The Next Decade in AI. their time is finally coming. Demis Hassabis on the limit in today’s AI: language can describe the world,

【採用情報】「Software Engineer」の5ポジションが現在オープン! https://sakana.ai/careers 「AIが進化すれば、ソフトウェアエンジニアの仕事はなくなるのか?」 Sakana AIは、全く逆だと考えています。 AIツールの登場で開発効率が劇…

ApplicationsDGX agent

【採用情報】「Software Engineer」の5ポジションが現在オープン! https://sakana.ai/careers 「AIが進化すれば、ソフトウェアエンジニアの仕事はなくなるのか?」 Sakana AIは、全く逆だと考えています。 AIツールの登場で開発効率が劇的に向上する一方、ジェボンズのパラドックス(Jevons paradox) が示すように、私たちが解決できる課題の幅と規

Symbolic Density Estimation for Discrete Distributions

Model ReleasesDGX agent

arXiv:2605.21813v1 Announce Type: new Abstract: Discrete probability laws underpin statistical modeling, yet the catalog of interpretable distributions has expanded only gradually through centuries of

Temporal Contrastive Transformer for Financial Crime Detection: Self-Supervised Sequence Embeddings via Predictive Contrastive Coding

ApplicationsDGX agent

arXiv:2605.21490v1 Announce Type: new Abstract: We introduce the Temporal Contrastive Transformer (TCT), a representation learning framework designed to capture contextual temporal dynamics in sequenc

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally ever…

Model ReleasesDGX agent

The case against me below is completely intellectually dishonest, filled with lies and misrepresentations, wrong about almost literally everything it says—a textbook example of propaganda: - I didn’t

the good old days, back when OpenAI was only losing $5 billion a year.

SafetyDGX agent

Gary Marcus comments on OpenAI's financial losses, referencing a period when the company's annual losses were approximately $5 billion as a point of comparison to its current financial situation. The

Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models

Model ReleasesDGX agent

Nemotron-Labs Diffusion Language Models represent NVIDIA's approach to achieving faster text generation through diffusion-based architectures, potentially offering significant speed improvements over

22 May 2026

A hacker group is poisoning open source code at an unprecedented scale

IndustryDGX agent

Cybercriminal group TeamPCP is executing an unprecedented wave of software supply chain attacks, compromising hundreds of open source tools and breaching organizations through poisoned code. The incid

Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development

AgentsDGX agent

arXiv:2605.20456v1 Announce Type: cross Abstract: Agentic AI coding systems can inspect repositories, plan implementation steps, edit files, call tools, run tests, and submit pull requests. These capa

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

Model ReleasesDGX agent

arXiv:2605.22366v1 Announce Type: new Abstract: Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However,

AI put 'synthetic quotes' in his book. But this author wants to keep using it.

IndustryDGX agent

Author Steven Rosenbaum's book 'The Future of Truth: How AI Reshapes Reality' was discovered by The New York Times to contain multiple misattributed or synthetic quotes generated by AI tools, which Ro

Artificial Intelligence Reshapes Microwave Photonics

AgentsDGX agent

arXiv:2605.21224v1 Announce Type: cross Abstract: As a rapidly emerging interdisciplinary field that intrinsically integrates microwave and photonics, microwave photonics (MWP) provides disruptive sol

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

AgentsDGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

Model ReleasesDGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

Eight key points from the most recent essay in the “AI as Normal Technology” series by @sayashk and me. Do AI Risks Require Extraordinary Go…

SafetyDGX agent

Eight key points from the most recent essay in the “AI as Normal Technology” series by @sayashk and me. Do AI Risks Require Extraordinary Government Intervention? 1. There is general consensus that AI

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

AgentsDGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

Evaluating multimodal emotion recognition in proactive conversational agents: A user study

AgentsDGX agent

arXiv:2605.20200v1 Announce Type: cross Abstract: This article presents a multimodal emotion recognition module integrated into a proactive Socially Interactive Agent (SIA) powered by generative artif

How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing

Model ReleasesDGX agent

arXiv:2602.01851v2 Announce Type: replace Abstract: Recent generative models have achieved remarkable progress in image editing. However, existing systems and benchmarks remain largely text-guided. In

HyperBench: Standardizing and Scaling Synthetic Evaluation for Hyperspectral Super-Resolution

ApplicationsDGX agent

arXiv:2605.21671v1 Announce Type: cross Abstract: Hyperspectral super-resolution (HSR) reconstructs a high-spatial-resolution hyperspectral image by fusing a low-resolution hyperspectral image (LR-HSI

i need someone at @OpenAI and @AnthropicAI to teach the models that while prototyping, backwards compatibility is just a bad idea

TutorialsDGX agent

Jeremy Howard argues that AI model developers at OpenAI and Anthropic should prioritize breaking backwards compatibility during the prototyping phase rather than maintaining it, suggesting that backwa

ImProver: Agent-Based Automated Proof Optimization

AgentsDGX agent

arXiv:2410.04753v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been used to generate formal proofs of mathematical theorems in proofs assistants such as Lean. However, we

InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation

Model ReleasesDGX agent

arXiv:2510.09724v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable of generating complete applications from natural language instructions, creating new opp

Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora

SafetyDGX agent

arXiv:2605.22660v1 Announce Type: new Abstract: Moral language is subtle and culturally variable, making it difficult to translate faithfully across languages. Idiomatic expressions, slang, and cultur

MTR-Bench: A Comprehensive Benchmark for Multi-Turn Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2505.17123v3 Announce Type: replace Abstract: Recent advances in Large Language Models (LLMs) have shown promising results in complex reasoning tasks. However, current evaluations predominantly

Open-World Evaluations for Measuring Frontier AI Capabilities

Model ReleasesDGX agent

arXiv:2605.20520v1 Announce Type: new Abstract: Benchmark-based evaluation remains important for tracking frontier AI progress. But it can both overstate and understate deployed capability because it

Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs

Model ReleasesDGX agent

arXiv:2605.22654v1 Announce Type: new Abstract: Previous detection studies have shown that LLMs cannot be effectively used as detectors, but these studies have not addressed modern Chinese poetry. Mor

SURGE: An Event-Centric Social Media Sentiment Time Series Benchmark with Interaction Structure

Model ReleasesDGX agent

arXiv:2605.21198v1 Announce Type: cross Abstract: Public events on social media generate large volumes of discussion whose collective dynamics carry direct value for opinion forecasting and crisis res

The Erdős Proof and AI Capabilities

SafetyDGX agent

View the official memo here. An internal model at OpenAI has autonomously disproved a central conjecture in discrete geometry, a mathematical field with applications in cryptography, wireless device c

Understanding Data Temporality Impact on Large Language Models Pre-training

Model ReleasesDGX agent

arXiv:2605.22769v1 Announce Type: new Abstract: Large language models (LLMs) are typically trained on shuffled corpora, yielding models whose knowledge is frozen at train time and whose temporal groun

VDE Bench: Evaluating The Capability of Image Editing Models to Modify Visual Documents

Model ReleasesDGX agent

arXiv:2602.00122v2 Announce Type: replace Abstract: In recent years, image editing models have made significant progress, enabling users to manipulate visual content in a flexible and interactive mann

We’re taking suggestions on what you want to see next week ✍️

Model ReleasesDGX agent

OpenAI solicited community feedback on X regarding content or features they should prioritize in the following week. This post reflects OpenAI's practice of engaging their audience to guide product de

21 May 2026

AI interoperability and layered trust emerge as the real unlocks for enterprise scale

ApplicationsDGX agent

Enterprise AI governance is becoming increasingly important as organizations race toward ROI, demanding AI systems that are scalable, predictable and built to deliver measurable business outcomes. Wit

ArchSIBench: Benchmarking the Architectural Spatial Intelligence of Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.20837v1 Announce Type: new Abstract: Architectural spatial intelligence, the ability to recognize and infer architectural space, is fundamental to tasks such as robot navigation, embodied i

Automated ICD Classification of Psychiatric Diagnoses: From Classical NLP to Large Language Models

Model ReleasesDGX agent

arXiv:2605.21154v1 Announce Type: new Abstract: Mental health has become a global priority, leading to a massive administrative burden in the coding of clinical diagnoses. This study proposes the auto

Can Vision Models Truly Forget? Mirage: Representation-Level Certification of Visual Unlearning

SafetyDGX agent

arXiv:2605.20282v1 Announce Type: new Abstract: Machine unlearning in Vertical Federated Learning (VFL) has attracted growing interest, yet existing methods certify forgetting solely using output-leve

CoarseSoundNet: Building a reliable model for ecological soundscape analysis

ApplicationsDGX agent

arXiv:2605.21143v1 Announce Type: cross Abstract: A soundscape is composed of three types of sound: biophony (sounds made by animals), geophony (natural abiotic sounds) and anthropophony (sounds made

Comparative Analysis of Military Detection Using Drone Imagery Across Multiple Visual Spectrums

ApplicationsDGX agent

arXiv:2605.21157v1 Announce Type: new Abstract: In modern warfare, drones are becoming an essential part of intelligence gathering and carrying out precise attacks in different kinds of hostile enviro

DarkShake-DVS: Event-based Human Action Recognition under Low-light andShaking Camera Conditions

Model ReleasesDGX agent

arXiv:2605.20680v1 Announce Type: new Abstract: Human Action Recognition (HAR) is a fundamental computer vision task with diverse real-world applications. Practical deployments often involve low-light

Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning

SafetyDGX agent

arXiv:2605.20730v1 Announce Type: new Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks through demonstrations, yet it suffers from escalating inference cos

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

SafetyDGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

← Previous
1…400401402403404…428
Next →