AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
20 May 2026

Smooth Piecewise Cutting for Neural Operator to Handle Discontinuities and Sharp Transitions

Model ReleasesDGX agent

arXiv:2605.19823v1 Announce Type: cross Abstract: Neural operators have achieved strong performance in learning solution operators of partial differential equations (PDEs), but their inherently contin

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

ApplicationsDGX agent

arXiv:2505.23747v2 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced performance on 2D visual tasks. However, improving

Tail Annealing for Heavy-Tailed Flow Matching

Model ReleasesDGX agent

arXiv:2605.20068v1 Announce Type: cross Abstract: Standard generative models struggle with heavy-tailed data: Lipschitz architectures cannot produce power-law tails from Gaussian noise, and interpolat

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

SafetyDGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

Model ReleasesDGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measur…

Model ReleasesDGX agent

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measuring every biomarker, or @sytses openly sharing and analyzing

The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility

Model ReleasesDGX agent

arXiv:2605.19537v1 Announce Type: new Abstract: Progress in LLMs is increasingly measured through standardized benchmarks, where state-of-the-art improvements are often separated by fractions of a per

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

Model ReleasesDGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

Model ReleasesDGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

What Do Evolutionary Coding Agents Evolve?

Model ReleasesDGX agent

arXiv:2605.20086v1 Announce Type: cross Abstract: Recent work pairs LLMs with evolutionary search to iteratively generate, modify, and select code using task-specific feedback. These systems have prod

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection

Model ReleasesDGX agent

arXiv:2601.22569v2 Announce Type: replace-cross Abstract: Large language model (LLM) based agents are increasingly used to automate financial transactions, yet their reliance on contextual reasoning e

19 May 2026

A Feature-Driven Framework for Software Fault Prediction

Model ReleasesDGX agent

arXiv:2605.17611v1 Announce Type: cross Abstract: Software fault prediction (SFP) is a critical task in software engineering, enabling early identification of faults in modules to improve software qua

A Machine With Human-Like Memory Systems

Model ReleasesDGX agent

arXiv:2204.01611v3 Announce Type: replace Abstract: Inspired by the cognitive science theory, we explicitly model an agent with both semantic and episodic memory systems, and show that it is better th

A Theory of Training Profit-Optimal LLMs

ResearchDGX agent

arXiv:2605.16430v1 Announce Type: cross Abstract: Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure

AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training

Local AiDGX agent

arXiv:2605.17923v1 Announce Type: cross Abstract: In video generation models, particularly world models, training large-scale video diffusion Transformers (such as DiT and MMDiT) poses significant com

Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory

Model ReleasesDGX agent

arXiv:2605.18733v1 Announce Type: new Abstract: Autoregressive video generation has improved rapidly in visual fidelity and interactivity, but it still suffers from long-term inconsistency and memory

AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent

AgentsDGX agent

arXiv:2602.03955v2 Announce Type: replace Abstract: While large language model (LLM) multi-agent systems achieve superior reasoning performance through iterative debate, practical deployment is limite

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

Model ReleasesDGX agent

arXiv:2605.17583v1 Announce Type: new Abstract: While existing text-to-speech (TTS) models exhibit high expressiveness, fine-grained control over composite instructions remains challenging due to the

AgentWall: A Runtime Safety Layer for Local AI Agents

Model ReleasesDGX agent

arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As agents transition from passive text generators to active ac

AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing

Model ReleasesDGX agent

arXiv:2603.23069v2 Announce Type: replace-cross Abstract: The task of authorship style transfer involves rewriting text in the style of a target author while preserving the meaning of the original tex

Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity

Model ReleasesDGX agent

arXiv:2510.00304v3 Announce Type: replace-cross Abstract: Deep learning models excel in stationary data but struggle in non-stationary environments due to a phenomenon known as loss of plasticity (LoP

Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs

Model ReleasesDGX agent

arXiv:2602.09805v2 Announce Type: replace-cross Abstract: As reasoning LLMs increasingly trade tokens for accuracy through deliberation, search, and self-correction, a single accuracy score can no lon

BioProAgent: Neuro-Symbolic Grounding for Constrained Scientific Planning

Model ReleasesDGX agent

arXiv:2603.00876v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant reasoning capabilities in scientific discovery but struggle to bridge the gap to physical

CAB: Accelerating Flow and Diffusion Sampling via Rectification and Corrected Adams-Bashforth

ResearchDGX agent

arXiv:2605.16736v1 Announce Type: new Abstract: Flow and diffusion models achieve high-fidelity, high-resolution image synthesis, but often require many function evaluations (NFEs) at sampling time. E

Can’t wait for Gemini Omni in @NotebookLM cinematic explainer videos 👀

Model ReleasesDGX agent

Emad Mostaque expressed anticipation for the integration of Google's Gemini Omni multimodal AI model into NotebookLM's cinematic explainer video generation features. The post suggests potential upcomi

CAREBench: Evaluating LLMs' Emotion Understanding by Assessing Cognitive Appraisal Reasoning

Model ReleasesDGX agent

arXiv:2605.17176v1 Announce Type: new Abstract: Emotion understanding is a core capability for LLMs to interact effectively with humans, yet existing evaluation paradigms rely on discrete emotion labe

ClawArena: Benchmarking AI Agents in Evolving Information Environments

Model ReleasesDGX agent

arXiv:2604.04202v2 Announce Type: replace-cross Abstract: AI agents deployed as persistent assistants must maintain correct beliefs as their information environment evolves. In practice, evidence is s

CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selection

Model ReleasesDGX agent

arXiv:2605.16839v1 Announce Type: new Abstract: Chunked prefill has become a widely adopted serving strategy for long-context large language models, but efficient attention computation in this regime

ContraFix: Agentic Vulnerability Repair via Differential Runtime Evidence and Skill Reuse

Model ReleasesDGX agent

arXiv:2605.17450v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used for automated vulnerability repair (AVR), where repository-level reasoning enables them to ins

DECODE: Domain-aware Continual Domain Expansion for Motion Prediction

AgentsDGX agent

arXiv:2411.17917v2 Announce Type: replace Abstract: Motion prediction is critical for autonomous vehicles to effectively navigate complex environments and accurately anticipate the behaviors of other

Diagnosing Korean-Language LLM Political Bias via Census-Grounded Agent Simulation

Local AiDGX agent

arXiv:2605.18395v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic political biases in voter simulations, but their underlying mechanisms and cross-lingual generalizatio

Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations

Model ReleasesDGX agent

arXiv:2605.17107v1 Announce Type: cross Abstract: We introduce a novel framework for uncertainty quantification of solution operators associated with stochastic partial differential equations (SPDEs).

DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes

Model ReleasesDGX agent

arXiv:2601.13839v2 Announce Type: replace Abstract: Social media imagery provides a low-latency source of situational information during natural and human-induced disasters, enabling rapid damage asse

DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2405.19189v3 Announce Type: replace Abstract: With the great success of diffusion models (DMs) in generating realistic synthetic vision data, many researchers have investigated their potential i

Evaluating Cognitive Age Alignment in Interactive AI Agents

Model ReleasesDGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Model ReleasesDGX agent

arXiv:2511.20857v2 Announce Type: replace-cross Abstract: Statefulness is essential for large language model (LLM) agents to perform long-term planning and problem-solving. This makes memory a critica

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective

Model ReleasesDGX agent

arXiv:2605.18421v1 Announce Type: cross Abstract: Recent benchmarks for Large Language Model (LLM) agents mainly evaluate reasoning, planning, and execution. However, memory is also essential for agen

Experimentally validated quantum-secure federated learning over a multi-user quantum network

Model ReleasesDGX agent

arXiv:2501.12709v2 Announce Type: replace-cross Abstract: Federated learning enables decentralized, privacy-preserving training but remains vulnerable to privacy leakage in the quantum era. Quantum fe

exttt{SynC}: Synergistic Boosting of Structure and Representation for Deep Graph Clustering

Model ReleasesDGX agent

arXiv:2406.15797v2 Announce Type: replace-cross Abstract: Employing graph neural networks (GNNs) for graph clustering has shown promising results in deep graph clustering. However, existing methods di

Fine-grained List-wise Alignment for Generative Medication Recommendation

Model ReleasesDGX agent

arXiv:2505.20218v2 Announce Type: replace Abstract: Accurate and safe medication recommendations are critical for effective clinical decision-making, especially in multimorbidity cases. However, exist

Flowing with Confidence

ResearchDGX agent

arXiv:2605.18472v1 Announce Type: cross Abstract: Generative models can produce nonsensical text, unrealistic images, and unstable materials faster than simulation or human review can absorb; without

From Documents to Segments: A Contextual Reformulation for Topic Assignment

ApplicationsDGX agent

arXiv:2605.17714v1 Announce Type: new Abstract: Traditional topic modeling assigns a single topic to each document. In practice, however, many real-world documents, such as product reviews or open-end

Gemini 3.5 Flash: more expensive, but Google plan to use it for everything

Model ReleasesDGX agent

Today at Google I/O, Google released Gemini 3.5 Flash. This one skipped the -preview modifier and went straight to general availability, and Google appear to be using it for a whole lot of their key p

Generative Artificial Intelligence for Literature Reviews

Model ReleasesDGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

Geometry-Aware Uncertainty Coresets for Robust Visual In-Context Learning in Histopathology

Model ReleasesDGX agent

arXiv:2605.18419v1 Announce Type: cross Abstract: Vision-language models (VLMs) can couple visual perception with open-ended clinical reasoning, making them attractive for computational histopathology

GRAFT: Decoupling Ranking and Calibration for Survival Analysis

ResearchDGX agent

arXiv:2602.07884v2 Announce Type: replace-cross Abstract: Survival analysis is complicated by censored data, high-dimensional features, and non-linear interactions. Classical models offer interpretabi

I/O 2026

Model ReleasesDGX agent

Google I/O 2026 is where Google shared how it's making AI more helpful for everyone, releasing new models including Gemini Omni and Gemini 3.5. The event showcased advancements to Google's agent-first

I/O 2026: Welcome to the agentic Gemini era

Model ReleasesDGX agent

At I/O 2026, Google announced that AI is transitioning from something users actively open to a background service that completes tasks automatically. The company introduced Gemini Spark, a new agentic

Language Game: Talking to Non-Human Systems

Model ReleasesDGX agent

arXiv:2605.16321v1 Announce Type: new Abstract: Language carries thought and coordination among humans but rarely reaches further along the spectrum of diverse intelligence. Yet non-neural systems --

Learning How to Cube

Model ReleasesDGX agent

arXiv:2605.16632v1 Announce Type: cross Abstract: Despite the effectiveness of Cube-and-Conquer (C&C) for solving challenging Boolean Satisfiability (SAT) problems, no prior work has shown that transf

LERA: LLM-Enhanced RAG for Ad Auction in Generative Chatbots

Model ReleasesDGX agent

arXiv:2605.16474v1 Announce Type: cross Abstract: The integration of advertising auction mechanisms into large language model (LLM)-based chatbots presents a significant opportunity for commercializat

Lightweight CNN-Based DDoS Detection for Resource-Constrained Edge Networks

Model ReleasesDGX agent

arXiv:2309.05646v2 Announce Type: replace-cross Abstract: Distributed Denial of Service (DDoS) attacks remain a persistent threat to the availability of Internet services, edge networks, and cyber-phy

llm-gemini 0.32

Model ReleasesDGX agent

llm-gemini 0.32 is an alpha release of Simon Willison's LLM Python library and CLI tool that provides access to Google's Gemini models , continuing work on major architectural changes to support newer

Lying with Truths: Open-Channel Multi-Agent Collusion for Belief Manipulation via Generative Montage

AgentsDGX agent

arXiv:2601.01685v2 Announce Type: replace-cross Abstract: As large language models (LLMs) transition to autonomous agents synthesizing real-time information, their reasoning capabilities introduce an

MedMIX: Modality-Internal Expert Fusion for Multimodal Medical Diagnosis

ResearchDGX agent

arXiv:2605.16639v1 Announce Type: new Abstract: Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing mod

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

Model ReleasesDGX agent

arXiv:2601.08118v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as human simulators, both for evaluating conversational systems and for generating fine-tuning da

Multilingual jailbreaking of LLMs using low-resource languages

Model ReleasesDGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

Neuroscience-inspired Staged Representation Learning with Disentangled Coarse- and Fine-Grained Semantics for EEG Visual Decoding

Model ReleasesDGX agent

arXiv:2605.16923v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) signals remains a fundamental challenge in brain-computer interfaces and medical rehabilit

NEWTON: Agentic Planning for Physically Grounded Video Generation

SafetyDGX agent

arXiv:2605.18396v1 Announce Type: new Abstract: Video generation models produce visually compelling results but systematically violate physical commonsense -- on VideoPhy-2, the best model achieves on

← Previous
1…411412413414415…1059
Next →