AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
Model Releases

EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding

DGX agent

arXiv:2605.09874v1 Announce Type: cross Abstract: Next-generation visual assistants, such as smart glasses, embodied agents, and always-on life-logging systems, must reason over an entire day or more

model-releasesarxiv-cs-ai
12 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

EmbodiSkill: Skill-Aware Reflection for Self-Evolving Embodied Agents

DGX agent

arXiv:2605.10332v1 Announce Type: new Abstract: Embodied agents can benefit from skills that guide object search, action execution, and state changes across diverse environments. Since embodied enviro

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EmoS: A High-Fidelity Multimodal Benchmark for Fine-grained Streaming Emotional Understanding

DGX agent

arXiv:2605.08847v1 Announce Type: new Abstract: In the context of today's high-pressure, aging society, the demand for large-scale emotional models capable of providing empathetic support is more crit

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents

DGX agent

arXiv:2605.09826v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to track others epistemic state, makes humans efficient collaborators. AI agents need the same capacity in multi agent

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

EnergyLens: Interpretable Closed-Form Energy Models for Multimodal LLM Inference Serving

DGX agent

arXiv:2605.10556v1 Announce Type: new Abstract: As large language models span dense, mixture-of-experts, and state-space architectures and are deployed on heterogeneous accelerators under increasingly

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

EpiGraph: A Knowledge Graph and Benchmark for Evidence-Intensive Reasoning in Epilepsy

DGX agent

arXiv:2605.09505v1 Announce Type: new Abstract: Epilepsy diagnosis and treatment require evidence-intensive reasoning across heterogeneous clinical knowledge, including biosignal patterns, genetic mec

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models

DGX agent

arXiv:2605.10410v1 Announce Type: new Abstract: Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

ER-Reason: A Benchmark Dataset for LLM Clinical Reasoning in the Emergency Room

DGX agent

arXiv:2505.22919v3 Announce Type: replace Abstract: Existing benchmarks for evaluating the clinical reasoning capabilities of large language models (LLMs) often lack a clear definition of 'clinical re

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

ERIS: Enhancing Privacy and Scalability in Federated Learning via Federated Shard Aggregation

DGX agent

arXiv:2602.08617v2 Announce Type: replace Abstract: Scaling Federated Learning (FL) to billion-parameter models forces a challenging trade-off between privacy, scalability, and model utility. Existing

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Even @haider1 sees that Mythos has been overhyped.

DGX agent

Even @haider1 sees that Mythos has been overhyped. mythos is pretty on par with gpt-5.5 and while gpt-5.5 is currently SOTA, it's not anything like what anthropic describes mythos as it's pretty obvio

model-releasesgary-marcus--x
12 May 2026
Model Releases

EverydayMMQA: A Multilingual and Multimodal Framework for Culturally Grounded Spoken Visual QA

DGX agent

arXiv:2510.06371v2 Announce Type: replace-cross Abstract: Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries re

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Exactness Matters for Physical Rule Enforcement

DGX agent

arXiv:2605.08285v1 Announce Type: new Abstract: Autoregressive scientific forecasters often enforce physical or structural constraints by repairing each predicted state before feeding it back into the

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Explanation Fairness in Large Language Models: An Empirical Analysis of Disparities in How LLMs Justify Decisions Across Demographic Groups

DGX agent

arXiv:2605.08671v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed not only to make decisions but to explain them. While AI decision fairness has been studied ext

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Explicit Reasoning Makes Better Judges: A Systematic Study on Accuracy, Efficiency, and Robustness

DGX agent

arXiv:2509.13332v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly adopted as automated judges in benchmarking and reward modeling, ensuring their reliability, effici

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models

DGX agent

arXiv:2605.09773v1 Announce Type: cross Abstract: We use sparse autoencoder (SAE) feature steering to amplify Dark Triad personality traits (Machiavellianism, narcissism, and psychopathy) in Llama-3.3

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk?

DGX agent

arXiv:2603.00166v2 Announce Type: replace-cross Abstract: Recent advances in generative AI have shown human-level performance in complex content creation. However, we identify a 'Paradox of Simplicity

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

expo: Exploration-prioritized policy optimization via adaptive kl regulation and gaussian curriculum sampling

DGX agent

arXiv:2605.09923v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the standard paradigm for LLM mathematical reasoning, where Group Relative Policy Optim

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FactoryNet: A Large-Scale Dataset toward Industrial Time-Series Foundation Models

DGX agent

arXiv:2605.09081v1 Announce Type: cross Abstract: We introduce the first universal pretraining corpus for industrial time-series data: FactoryNet. 51M datapoints across 23k end-to-end task executions

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Failing Forward: Adaptive Failure-Informed Learning for Vision-Language-Action Models

DGX agent

arXiv:2605.08434v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising paradigm for scalable robotic manipulation, yet their reliance on success-only behavioral clonin

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Falcon 9 launches NROL-172 to orbit from pad 4E in California

DGX agent

SpaceX's Falcon 9 rocket successfully launched the NROL-172 classified national reconnaissance payload from Space Launch Complex 4E at Vandenberg Space Force Base in California. This mission was condu

model-releaseselon-musk--x
12 May 2026
Model Releases

Fashion Florence: Fine-Tuning Florence-2 for Structured Fashion Attribute Extraction

DGX agent

arXiv:2605.09827v1 Announce Type: cross Abstract: We present Fashion Florence, a Florence-2 vision-language model fine-tuned with LoRA to extract structured fashion attributes from clothing images. Gi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition

DGX agent

arXiv:2605.10127v1 Announce Type: new Abstract: Recent research work on fashion outfit generation focuses on promoting visual consistency of garments by leveraging key information from reference image

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Fast mode for Claude Opus 4.7 is now available in Cursor! It's 2.5x the speed at 6x the cost. For most tasks, we recommend using the standar…

DGX agent

Cursor has released fast mode for Claude Opus 4.7, offering 2.5x faster processing speeds but at 6x the cost compared to standard mode. The announcement suggests that for most tasks, standard mode rem

model-releasescursor--x
12 May 2026
Model Releases

Fast mode for Claude Opus 4.7 is now available in research preview on the API and in Claude Code.

DGX agent

Anthropic has released a fast mode for Claude Opus 4.7, now available in research preview for both the API and Claude Code, offering improved performance for compatible workloads. This feature allows

model-releasesboris-cherny--x
12 May 2026
Model Releases

Feature Repulsion and Spectral Lock-in: An Empirical Study of Two-Layer Network Grokking

DGX agent

arXiv:2605.08119v1 Announce Type: cross Abstract: Tian (2025) proves a repulsion theorem (Theorem 6) for the matrix B = (widetilde{F}^op widetilde{F} + eta I)^{-1} during the interactive feature-learn

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Feature Rivalry in Sparse Autoencoder Representations: A Mechanistic Study of Uncertainty-Driven Feature Competition in LLMs

DGX agent

arXiv:2605.08149v1 Announce Type: cross Abstract: Sparse Autoencoders (SAEs) decompose large language model representations into interpretable features, but how these features interact under uncertain

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Featurized Occupation Measures for Structured Global Search in Numerical Optimal Control

DGX agent

arXiv:2603.16231v2 Announce Type: replace-cross Abstract: Numerical optimal control has long been split between globally structured but dimensionally intractable Hamilton--Jacobi--Bellman (HJB) method

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

Federated Language Models Under Bandwidth Budgets: Distillation Rates and Conformal Coverage

DGX agent

arXiv:2605.09986v1 Announce Type: cross Abstract: Training a language model on data scattered across bandwidth-limited nodes that cannot be centralized is a setting that arises in clinical networks, e

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Filtering Memorization from Parameter-Space in Diffusion Models

DGX agent

arXiv:2605.10439v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely used mechanism for customizing diffusion models, enabling users to inject new visual concepts or styles t

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Fin-Bias: Comprehensive Evaluation for LLM Decision-Making under human bias in Finance Domain

DGX agent

arXiv:2605.09106v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in financial contexts, raising critical concerns about reliability, alignment, and susceptibility

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

FinTSB: A Comprehensive and Practical Benchmark for Financial Time Series Forecasting

DGX agent

arXiv:2502.18834v2 Announce Type: replace-cross Abstract: Financial time series (FinTS) record the behavior of human-brain-augmented decision-making, capturing valuable historical information that can

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Fitting Multilinear Polynomials for Logic Gate Networks

DGX agent

arXiv:2605.08657v1 Announce Type: cross Abstract: We study learnable logic gate networks that stack layers of 2-input Boolean gates to build combinational circuits. Every 2-input gate has a unique mul

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Fix the Loss, Not the Radius: Rethinking the Adversarial Perturbation of Sharpness-Aware Minimization

DGX agent

arXiv:2605.10183v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) improves generalization by minimizing the worst-case loss within a fixed parameter-space radius neighborhood. SAM and

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning

DGX agent

arXiv:2605.09355v1 Announce Type: new Abstract: Real-world model deployment across multiple domains requires multimodal models to operate under two complementary regimes: (1) multi-task pretraining, t

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

DGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FlashClear: Ultra-Fast Image Content Removal via Efficient Step Distillation and Feature Caching

DGX agent

arXiv:2605.09003v1 Announce Type: new Abstract: Recently, diffusion-based object removal models have achieved impressive results in eliminating objects and their associated visual effects. However, th

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Follow the Mean: Reference-Guided Flow Matching

DGX agent

arXiv:2605.10302v1 Announce Type: new Abstract: Existing approaches to controllable generation typically rely on fine-tuning, auxiliary networks, or test-time search. We show that flow matching admits

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs

DGX agent

arXiv:2605.08905v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable success on reasoning benchmarks through Reinforcement Learning with Verifiable Rewards (RLVR), exc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FormalRewardBench: A Benchmark for Formal Theorem Proving Reward Models

DGX agent

arXiv:2605.10141v1 Announce Type: new Abstract: Recent neural theorem provers use reinforcement learning with verifiable rewards (RLVR), where proof assistants provide binary correctness signals. Whil

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FORTIS: Benchmarking Over-Privilege in Agent Skills

DGX agent

arXiv:2605.09163v1 Announce Type: new Abstract: Large language model agents increasingly operate through an intermediate skill layer that mediates between user intent and concrete task execution. This

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling

DGX agent

arXiv:2512.02010v5 Announce Type: replace Abstract: As large language models have grown larger, interest has grown in low-precision numerical formats such as NVFP4 as a way to improve speed and reduce

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

FPGA-Based Hardware Architecture for Contrast Maximization in Event-Based Vision

DGX agent

arXiv:2605.09581v1 Announce Type: new Abstract: This paper presents a hardware architecture that implements the Contrast Maximization (CM) algorithm in Field-Programmable Gate Array (FPGA) resources f

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

FRACTAL: SSM with Fractional Recurrent Architecture for Computational Temporal Analysis of Long Sequences

DGX agent

arXiv:2605.08833v1 Announce Type: new Abstract: Effective sequence modeling fundamentally requires balancing the retention of unbounded history with the high-resolution detection of abrupt short-term

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Frame In, Frame Out: Measuring Framing Bias in LLM-Generated News Summaries

DGX agent

arXiv:2505.05406v2 Announce Type: replace Abstract: News headlines and summaries shape how events are interpreted through selective emphasis and omission, a phenomenon commonly referred to as framing.

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

FraudBench: A Multimodal Benchmark for Detecting AI-Generated Fraudulent Refund Evidence

DGX agent

arXiv:2605.08820v1 Announce Type: cross Abstract: Artificial Intelligence (AI)-generated images have become increasingly realistic and readily adaptable to concrete real-world claims, creating new cha

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

FreeMOCA: Memory-Free Continual Learning for Malicious Code Analysis

DGX agent

arXiv:2605.09664v1 Announce Type: cross Abstract: As over 200 million new malware samples are identified each year, antivirus systems must continuously adapt to the evolving threat landscape. However,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

From Articulated Kinematics to Routed Visual Control for Action-Conditioned Surgical Video Generation

DGX agent

arXiv:2605.08712v1 Announce Type: new Abstract: Action-conditioned surgical video generation is a critical yet highly challenging problem for robotic surgery. The core difficulty is that low-dimension

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

From Pixels to Concepts: Do Segmentation Models Understand What They Segment?

DGX agent

arXiv:2605.09591v1 Announce Type: new Abstract: Segmentation is a fundamental vision task underlying numerous downstream applications. Recent promptable segmentation models, such as Segment Anything M

model-releasesarxiv-cs-cv
12 May 2026
← Previous
1…327328329330331…471
Next →