AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,598
  • Agents7,796
  • Applications5,565
  • Concepts5
  • Hardware1,944
  • Industry6,220
  • Local Ai5,134
  • Model Releases24,972
  • Research20,928
  • Safety13,838
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,598Total entries
1Added by human
91,597Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,253 results
Model Releases

Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity

DGX agent

arXiv:2510.00304v3 Announce Type: replace-cross Abstract: Deep learning models excel in stationary data but struggle in non-stationary environments due to a phenomenon known as loss of plasticity (LoP

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs

DGX agent

arXiv:2602.09805v2 Announce Type: replace-cross Abstract: As reasoning LLMs increasingly trade tokens for accuracy through deliberation, search, and self-correction, a single accuracy score can no lon

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

BioProAgent: Neuro-Symbolic Grounding for Constrained Scientific Planning

DGX agent

arXiv:2603.00876v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant reasoning capabilities in scientific discovery but struggle to bridge the gap to physical

model-releasesarxiv-cs-ai
19 May 2026
Research

CAB: Accelerating Flow and Diffusion Sampling via Rectification and Corrected Adams-Bashforth

DGX agent

arXiv:2605.16736v1 Announce Type: new Abstract: Flow and diffusion models achieve high-fidelity, high-resolution image synthesis, but often require many function evaluations (NFEs) at sampling time. E

researcharxiv-cs-cv
19 May 2026
Model Releases

Can’t wait for Gemini Omni in @NotebookLM cinematic explainer videos 👀

DGX agent

Emad Mostaque expressed anticipation for the integration of Google's Gemini Omni multimodal AI model into NotebookLM's cinematic explainer video generation features. The post suggests potential upcomi

model-releasesemad-mostaque--x
19 May 2026
Model Releases

CAREBench: Evaluating LLMs' Emotion Understanding by Assessing Cognitive Appraisal Reasoning

DGX agent

arXiv:2605.17176v1 Announce Type: new Abstract: Emotion understanding is a core capability for LLMs to interact effectively with humans, yet existing evaluation paradigms rely on discrete emotion labe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ClawArena: Benchmarking AI Agents in Evolving Information Environments

DGX agent

arXiv:2604.04202v2 Announce Type: replace-cross Abstract: AI agents deployed as persistent assistants must maintain correct beliefs as their information environment evolves. In practice, evidence is s

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selection

DGX agent

arXiv:2605.16839v1 Announce Type: new Abstract: Chunked prefill has become a widely adopted serving strategy for long-context large language models, but efficient attention computation in this regime

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

ContraFix: Agentic Vulnerability Repair via Differential Runtime Evidence and Skill Reuse

DGX agent

arXiv:2605.17450v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used for automated vulnerability repair (AVR), where repository-level reasoning enables them to ins

model-releasesarxiv-cs-ai
19 May 2026
Agents

DECODE: Domain-aware Continual Domain Expansion for Motion Prediction

DGX agent

arXiv:2411.17917v2 Announce Type: replace Abstract: Motion prediction is critical for autonomous vehicles to effectively navigate complex environments and accurately anticipate the behaviors of other

agentsarxiv-cs-cv
19 May 2026
Local Ai

Diagnosing Korean-Language LLM Political Bias via Census-Grounded Agent Simulation

DGX agent

arXiv:2605.18395v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic political biases in voter simulations, but their underlying mechanisms and cross-lingual generalizatio

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations

DGX agent

arXiv:2605.17107v1 Announce Type: cross Abstract: We introduce a novel framework for uncertainty quantification of solution operators associated with stochastic partial differential equations (SPDEs).

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes

DGX agent

arXiv:2601.13839v2 Announce Type: replace Abstract: Social media imagery provides a low-latency source of situational information during natural and human-induced disasters, enabling rapid damage asse

model-releasesarxiv-cs-cv
19 May 2026
Safety

DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning

DGX agent

arXiv:2405.19189v3 Announce Type: replace Abstract: With the great success of diffusion models (DMs) in generating realistic synthetic vision data, many researchers have investigated their potential i

safetyarxiv-cs-lg
19 May 2026
Model Releases

Evaluating Cognitive Age Alignment in Interactive AI Agents

DGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

DGX agent

arXiv:2511.20857v2 Announce Type: replace-cross Abstract: Statefulness is essential for large language model (LLM) agents to perform long-term planning and problem-solving. This makes memory a critica

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective

DGX agent

arXiv:2605.18421v1 Announce Type: cross Abstract: Recent benchmarks for Large Language Model (LLM) agents mainly evaluate reasoning, planning, and execution. However, memory is also essential for agen

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Experimentally validated quantum-secure federated learning over a multi-user quantum network

DGX agent

arXiv:2501.12709v2 Announce Type: replace-cross Abstract: Federated learning enables decentralized, privacy-preserving training but remains vulnerable to privacy leakage in the quantum era. Quantum fe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

exttt{SynC}: Synergistic Boosting of Structure and Representation for Deep Graph Clustering

DGX agent

arXiv:2406.15797v2 Announce Type: replace-cross Abstract: Employing graph neural networks (GNNs) for graph clustering has shown promising results in deep graph clustering. However, existing methods di

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Fine-grained List-wise Alignment for Generative Medication Recommendation

DGX agent

arXiv:2505.20218v2 Announce Type: replace Abstract: Accurate and safe medication recommendations are critical for effective clinical decision-making, especially in multimorbidity cases. However, exist

model-releasesarxiv-cs-lg
19 May 2026
Research

Flowing with Confidence

DGX agent

arXiv:2605.18472v1 Announce Type: cross Abstract: Generative models can produce nonsensical text, unrealistic images, and unstable materials faster than simulation or human review can absorb; without

researcharxiv-cs-ai
19 May 2026
Applications

From Documents to Segments: A Contextual Reformulation for Topic Assignment

DGX agent

arXiv:2605.17714v1 Announce Type: new Abstract: Traditional topic modeling assigns a single topic to each document. In practice, however, many real-world documents, such as product reviews or open-end

applicationsarxiv-cs-cl
19 May 2026
Model Releases

Gemini 3.5 Flash: more expensive, but Google plan to use it for everything

DGX agent

Today at Google I/O, Google released Gemini 3.5 Flash. This one skipped the -preview modifier and went straight to general availability, and Google appear to be using it for a whole lot of their key p

model-releasessimon-willison
19 May 2026
Model Releases

Generative Artificial Intelligence for Literature Reviews

DGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Geometry-Aware Uncertainty Coresets for Robust Visual In-Context Learning in Histopathology

DGX agent

arXiv:2605.18419v1 Announce Type: cross Abstract: Vision-language models (VLMs) can couple visual perception with open-ended clinical reasoning, making them attractive for computational histopathology

model-releasesarxiv-cs-ai
19 May 2026
Research

GRAFT: Decoupling Ranking and Calibration for Survival Analysis

DGX agent

arXiv:2602.07884v2 Announce Type: replace-cross Abstract: Survival analysis is complicated by censored data, high-dimensional features, and non-linear interactions. Classical models offer interpretabi

researcharxiv-cs-ai
19 May 2026
Model Releases

I/O 2026

DGX agent

Google I/O 2026 is where Google shared how it's making AI more helpful for everyone, releasing new models including Gemini Omni and Gemini 3.5. The event showcased advancements to Google's agent-first

model-releasesgoogle-ai
19 May 2026
Model Releases

I/O 2026: Welcome to the agentic Gemini era

DGX agent

At I/O 2026, Google announced that AI is transitioning from something users actively open to a background service that completes tasks automatically. The company introduced Gemini Spark, a new agentic

model-releasesgoogle-ai
19 May 2026
Model Releases

Language Game: Talking to Non-Human Systems

DGX agent

arXiv:2605.16321v1 Announce Type: new Abstract: Language carries thought and coordination among humans but rarely reaches further along the spectrum of diverse intelligence. Yet non-neural systems --

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Learning How to Cube

DGX agent

arXiv:2605.16632v1 Announce Type: cross Abstract: Despite the effectiveness of Cube-and-Conquer (C&C) for solving challenging Boolean Satisfiability (SAT) problems, no prior work has shown that transf

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LERA: LLM-Enhanced RAG for Ad Auction in Generative Chatbots

DGX agent

arXiv:2605.16474v1 Announce Type: cross Abstract: The integration of advertising auction mechanisms into large language model (LLM)-based chatbots presents a significant opportunity for commercializat

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Lightweight CNN-Based DDoS Detection for Resource-Constrained Edge Networks

DGX agent

arXiv:2309.05646v2 Announce Type: replace-cross Abstract: Distributed Denial of Service (DDoS) attacks remain a persistent threat to the availability of Internet services, edge networks, and cyber-phy

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

llm-gemini 0.32

DGX agent

llm-gemini 0.32 is an alpha release of Simon Willison's LLM Python library and CLI tool that provides access to Google's Gemini models , continuing work on major architectural changes to support newer

model-releasessimon-willison
19 May 2026
Agents

Lying with Truths: Open-Channel Multi-Agent Collusion for Belief Manipulation via Generative Montage

DGX agent

arXiv:2601.01685v2 Announce Type: replace-cross Abstract: As large language models (LLMs) transition to autonomous agents synthesizing real-time information, their reasoning capabilities introduce an

agentsarxiv-cs-ai
19 May 2026
Research

MedMIX: Modality-Internal Expert Fusion for Multimodal Medical Diagnosis

DGX agent

arXiv:2605.16639v1 Announce Type: new Abstract: Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing mod

researcharxiv-cs-lg
19 May 2026
Model Releases

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

DGX agent

arXiv:2601.08118v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as human simulators, both for evaluating conversational systems and for generating fine-tuning da

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Multilingual jailbreaking of LLMs using low-resource languages

DGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Neuroscience-inspired Staged Representation Learning with Disentangled Coarse- and Fine-Grained Semantics for EEG Visual Decoding

DGX agent

arXiv:2605.16923v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) signals remains a fundamental challenge in brain-computer interfaces and medical rehabilit

model-releasesarxiv-cs-cv
19 May 2026
Safety

NEWTON: Agentic Planning for Physically Grounded Video Generation

DGX agent

arXiv:2605.18396v1 Announce Type: new Abstract: Video generation models produce visually compelling results but systematically violate physical commonsense -- on VideoPhy-2, the best model achieves on

safetyarxiv-cs-cv
19 May 2026
Research

Nonlinear Bipolar Compensation: Handling Outliers in Post-Training Quantization

DGX agent

arXiv:2605.16423v1 Announce Type: new Abstract: Network quantization has emerged as one of the most practical model compression techniques, which significantly reduces a model's memory and compute con

researcharxiv-cs-cv
19 May 2026
Model Releases

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

DGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

model-releasesarxiv-cs-ai
19 May 2026
Safety

Old Habits Die Hard: How Conversational History Geometrically Traps LLMs

DGX agent

arXiv:2603.03308v2 Announce Type: replace-cross Abstract: How does the conversational past of large language models (LLMs) influence their future performance? Recent work suggests that LLMs are affect

safetyarxiv-cs-ai
19 May 2026
Model Releases

open sourcing Marlin-2B 🐟 a tiny VLM to extract structured information from videos Marlin is finetuned for two questions devs want to ask i…

DGX agent

open sourcing Marlin-2B 🐟 a tiny VLM to extract structured information from videos Marlin is finetuned for two questions devs want to ask in their videos: what is happening, and when? Best open model

model-releasesclem-delangue--x
19 May 2026
Model Releases

Optimising CSRNet with parameter-free attention mechanisms for crowd counting in public transport

DGX agent

arXiv:2605.18349v1 Announce Type: cross Abstract: Occupancy estimation and crowd counting are critical tasks in designing smart and efficient public transport vehicles. Given that public transport loa

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Parameter-Efficient Domain Adaptation of Physics-Informed Self-Attention based GNNs for AC Power Flow Prediction

DGX agent

arXiv:2602.18227v2 Announce Type: replace Abstract: Accurate AC power flow (AC-PF) prediction under domain shift is critical when models trained on medium-voltage (MV) grids are deployed on high-volta

model-releasesarxiv-cs-lg
19 May 2026
Safety

Presupposition and Reasoning in Conditionals: A Theory-Based Study of Humans and LLMs

DGX agent

arXiv:2605.18352v1 Announce Type: new Abstract: Presupposition projection in conditionals is central to theories of meaning and pragmatics, yet it remains largely unevaluated in large language models.

safetyarxiv-cs-cl
19 May 2026
Model Releases

PRIME: Physically-consistent Robotic Inertial and Motion Estimation for Legged and Humanoid Robots

DGX agent

arXiv:2605.17681v1 Announce Type: new Abstract: Humanoid and legged robots interact with the environment through intermittent contacts, making accurate motion estimation fundamentally dependent on rea

model-releasesarxiv-cs-ro
19 May 2026
Local Ai

R2V Agent: Teaching SLMs When to Ask for Help

DGX agent

arXiv:2605.16604v1 Announce Type: new Abstract: Efficient agentic systems should incur expensive frontier-model costs only on decisions where a cheaper local model is likely to fail. Existing LLM casc

local-aiarxiv-cs-lg
19 May 2026
← Previous
1…539540541542543…1381
Next →