AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
20 Apr 2026

Video-STAR: Reinforcing Open-Vocabulary Action Recognition with Tools

ResearchDGX agent

arXiv:2510.08480v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated remarkable potential in bridging visual and textual reasoning, yet their reliance on text

19 Apr 2026

You should watch this. It just shows how disconnected we are from the small group of people making decisions that will impact our future hea…

Model ReleasesDGX agent

You should watch this. It just shows how disconnected we are from the small group of people making decisions that will impact our future heavily. These people have so much ai psychosis. If you listen

18 Apr 2026

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Airbnb launches a pilot in NYC, LA, and other cities that lets users to select from a range of boutique hotels alongside private homes in a bid to boost growth (Stephanie Stacey/Financial Times)
Model ReleasesDGX agent

Stephanie Stacey / Financial Times: Airbnb launches a pilot in NYC, LA, and other cities that lets users to select from a range of boutique hotels alongside private homes in a bid to boost growth — Co

Claude system prompts as a git timeline

Model ReleasesDGX agent

Research: Claude system prompts as a git timeline Anthropic publish the system prompts for Claude chat and make that page available as Markdown. I had Claude Code turn that page into separate files fo

My Workflow for Understanding LLM Architectures

ResearchDGX agent

This article outlines Sebastian Raschka's systematic approach to learning and understanding Large Language Model (LLM) architectures, likely covering foundational concepts, key components, and methodo

One of the premier journals in my field... I think there are very valid reasons to set rules on AI in peer review (including disclosure), bu…

ApplicationsDGX agent

One of the premier journals in my field... I think there are very valid reasons to set rules on AI in peer review (including disclosure), but the idea that all AI models steal your data is very 2023.

Salesforce announces Headless 360, an initiative that will give AI agents access to Salesforce's platform capabilities through APIs, MCP tools or CLI commands (Michael Nuñez/VentureBeat)

Model ReleasesDGX agent

Michael Nuñez / VentureBeat: Salesforce announces Headless 360, an initiative that will give AI agents access to Salesforce's platform capabilities through APIs, MCP tools or CLI commands — Salesforce

We push Prefill/Decode disaggregation beyond a single cluster: cross-datacenter + heterogeneous hardware, unlocking the potential for signif…

Model ReleasesDGX agent

We push Prefill/Decode disaggregation beyond a single cluster: cross-datacenter + heterogeneous hardware, unlocking the potential for significantly lower cost per token. This was previously blocked by

17 Apr 2026

A Mechanistic Account of Attention Sinks in GPT-2: One Circuit, Broader Implications for Mitigation

SafetyDGX agent

arXiv:2604.14722v1 Announce Type: new Abstract: Transformers commonly exhibit an attention sink: disproportionately high attention to the first position. We study this behavior in GPT-2-style models w

A multi-platform LiDAR dataset for standardized forest inventory measurement at long term ecological monitoring sites

Model ReleasesDGX agent

arXiv:2604.14635v1 Announce Type: new Abstract: We present a curated multi-platform LiDAR reference dataset from an instrumented ICOS forest plot, explicitly designed to support calibration, benchmark

A Nonlinear Separation Principle: Applications to Neural Networks, Control and Learning

Model ReleasesDGX agent

arXiv:2604.15238v1 Announce Type: cross Abstract: This paper investigates continuous-time and discrete-time firing-rate and Hopfield recurrent neural networks (RNNs), with applications in nonlinear co

ADAPT: Benchmarking Commonsense Planning under Unspecified Affordance Constraints

Model ReleasesDGX agent

arXiv:2604.14902v1 Announce Type: cross Abstract: Intelligent embodied agents should not simply follow instructions, as real-world environments often involve unexpected conditions and exceptions. Howe

AgentGA: Evolving Code Solutions in Agent-Seed Space

Model ReleasesDGX agent

arXiv:2604.14655v1 Announce Type: cross Abstract: We present AgentGA, a framework that evolves autonomous code-generation runs by optimizing the agent seed: the task prompt plus optional parent archiv

AgentIAD: Agentic Industrial Anomaly Detection via Adaptive Memory Augmentation

Model ReleasesDGX agent

arXiv:2512.13671v2 Announce Type: replace Abstract: Industrial anomaly detection (IAD) is challenging due to the subtle and highly localized nature of many defects, which single-pass vision--language

AIM: Asymmetric Information Masking for Visual Question Answering Continual Learning

ResearchDGX agent

arXiv:2604.14779v1 Announce Type: cross Abstract: In continual visual question answering (VQA), existing Continual Learning (CL) methods are mostly built for symmetric, unimodal architectures. However

Awakening Dormant Experts:Counterfactual Routing to Mitigate MoE Hallucinations

ResearchDGX agent

arXiv:2604.14246v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models have achieved remarkable scalability, yet they remain vulnerable to hallucinations, particularly when processing

Benchmarking Classical Coverage Path Planning Heuristics on Irregular Hexagonal Grids for Maritime Coverage Scenarios

Model ReleasesDGX agent

arXiv:2604.15202v1 Announce Type: new Abstract: Coverage path planning on irregular hexagonal grids is relevant to maritime surveillance, search and rescue and environmental monitoring, yet classical

Bird-SR: Bidirectional Reward-Guided Diffusion for Real-World Image Super-Resolution

SafetyDGX agent

arXiv:2602.07069v2 Announce Type: replace Abstract: Powered by multimodal text-to-image priors, diffusion-based super-resolution excels at synthesizing intricate details; however, models trained on sy

CausalEmbed: Auto-Regressive Multi-Vector Generation in Latent Space for Visual Document Embedding

TutorialsDGX agent

arXiv:2601.21262v3 Announce Type: replace Abstract: Although Multimodal Large Language Models (MLLMs) have shown remarkable potential in Visual Document Retrieval (VDR) through generating high-quality

CLion: Efficient Cautious Lion Optimizer with Enhanced Generalization

Model ReleasesDGX agent

arXiv:2604.14587v1 Announce Type: new Abstract: Lion optimizer is a popular learning-based optimization algorithm in machine learning, which shows impressive performance in training many deep learning

Domain Fine-Tuning FinBERT on Finnish Histopathological Reports: Train-Time Signals and Downstream Correlations

ApplicationsDGX agent

arXiv:2604.14815v1 Announce Type: new Abstract: In NLP classification tasks where little labeled data exists, domain fine-tuning of transformer models on unlabeled data is an established approach. In

ECM Contracts: Contract-Aware, Versioned, and Governable Capability Interfaces for Embodied Agents

Model ReleasesDGX agent

arXiv:2604.13097v1 Announce Type: cross Abstract: Embodied agents increasingly rely on modular capabilities that can be installed, upgraded, composed, and governed at runtime. Prior work has introduce

Energy-based Regularization for Learning Residual Dynamics in Neural MPC for Omnidirectional Aerial Robots

ApplicationsDGX agent

arXiv:2604.14678v1 Announce Type: cross Abstract: Data-driven Model Predictive Control (MPC) has lately been the core research subject in the field of control theory. The combination of an optimal con

Exploiting Correlations in Federated Learning: Opportunities and Practical Limitations

Model ReleasesDGX agent

arXiv:2604.14751v1 Announce Type: cross Abstract: The communication bottleneck in federated learning (FL) has spurred extensive research into techniques to reduce the volume of data exchanged between

FAIR Universe Weak Lensing ML Uncertainty Challenge: Handling Uncertainties and Distribution Shifts for Precision Cosmology

Model ReleasesDGX agent

arXiv:2604.14451v1 Announce Type: cross Abstract: Weak gravitational lensing, the correlated distortion of background galaxy shapes by foreground structures, is a powerful probe of the matter distribu

FedIDM: Achieving Fast and Stable Convergence in Byzantine Federated Learning through Iterative Distribution Matching

Model ReleasesDGX agent

arXiv:2604.15115v1 Announce Type: new Abstract: Most existing Byzantine-robust federated learning (FL) methods suffer from slow and unstable convergence. Moreover, when handling a substantial proporti

Flux Dev.1 Artistic Mix 04-16-2026

Local AiDGX agent

Flux Dev is a higher-quality variant of the Flux.1 image generation model that is non-commercial . The Reddit post titled 'Flux Dev.1 Artistic Mix 04-16-2026' likely showcases a community member's cur

FreqTrack: Frequency Learning based Vision Transformer for RGB-Event Object Tracking

Model ReleasesDGX agent

arXiv:2604.14526v1 Announce Type: new Abstract: Existing single-modal RGB trackers often face performance bottlenecks in complex dynamic scenes, while the introduction of event sensors offers new pote

Generalization in LLM Problem Solving: The Case of the Shortest Path

ResearchDGX agent

arXiv:2604.15306v1 Announce Type: cross Abstract: Whether language models can systematically generalize remains actively debated. Yet empirical performance is jointly shaped by multiple factors such a

H2VLR: Heterogeneous Hypergraph Vision-Language Reasoning for Few-Shot Anomaly Detection

Model ReleasesDGX agent

arXiv:2604.14507v1 Announce Type: new Abstract: As a classic vision task, anomaly detection has been widely applied in industrial inspection and medical imaging. In this task, data scarcity is often a

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding

HardwareDGX agent

arXiv:2601.14724v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated significant improvement in offline video understanding. Howe

Hybrid Latents -- Geometry-Appearance-Aware Surfel Splatting

SafetyDGX agent

arXiv:2604.14928v1 Announce Type: new Abstract: We introduce a hybrid Gaussian-hash-grid radiance representation for reconstructing 2D Gaussian scene models from multi-view images. Similar to NeST spl

IROSA: Interactive Robot Skill Adaptation using Natural Language

SafetyDGX agent

arXiv:2603.03897v3 Announce Type: replace-cross Abstract: Foundation models have demonstrated impressive capabilities across diverse domains, while imitation learning provides principled methods for r

Join us at PyCon US 2026 in Long Beach - we have new AI and security tracks this year

Model ReleasesDGX agent

This year's PyCon US is coming up next month from May 13th to May 19th, with the core conference talks from Friday 15th to Sunday 17th and tutorial and sprint days either side. It's in Long Beach, Cal

Learning to Think Like a Cartoon Captionist: Incongruity-Resolution Supervision for Multimodal Humor Understanding

SafetyDGX agent

arXiv:2604.15210v1 Announce Type: cross Abstract: Humor is one of the few cognitive tasks where getting the reasoning right matters as much as getting the answer right. While recent work evaluates hum

LLM agents loop, drift, and get stuck on hard reasoning tasks up to 30% of the time. Current fixes are either too blunt (hard step limits) o…

TutorialsDGX agent

LLM agents loop, drift, and get stuck on hard reasoning tasks up to 30% of the time. Current fixes are either too blunt (hard step limits) or too expensive (LLM-as-judge adding 10-15% overhead per ste

MIOFlow 2.0: A unified framework for inferring cellular stochastic dynamics from single cell and spatial transcriptomics data

ResearchDGX agent

arXiv:2603.22564v2 Announce Type: replace Abstract: Understanding cellular trajectories via time-resolved single-cell transcriptomics is vital for studying development, regeneration, and disease. A ke

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining

Model ReleasesDGX agent

arXiv:2604.14198v1 Announce Type: cross Abstract: Domain reweighting can improve sample efficiency and downstream generalization, but data-mixture optimization for multimodal midtraining remains large

OmniCompliance-100K: A Multi-Domain, Rule-Grounded, Real-World Safety Compliance Dataset

SafetyDGX agent

arXiv:2603.13933v2 Announce Type: replace Abstract: Ensuring the safety and compliance of large language models (LLMs) is of paramount importance. However, existing LLM safety datasets often rely on a

Physics-Informed Machine Learning for Pouch Cell Temperature Estimation

ResearchDGX agent

arXiv:2604.14566v1 Announce Type: new Abstract: Accurate temperature estimation of pouch cells with indirect liquid cooling is essential for optimizing battery thermal management systems for transport

PolyBench: Benchmarking LLM Forecasting and Trading Capabilities on Live Prediction Market Data

Model ReleasesDGX agent

arXiv:2604.14199v1 Announce Type: cross Abstract: Predicting real-world events from live market signals demands systems that fuse qualitative news with quantitative order-book dynamics under strict te

Portfolio Optimization Proxies under Label Scarcity and Regime Shifts via Bayesian and Deterministic Students under Semi-Supervised Sandwich Training

ResearchDGX agent

arXiv:2604.14206v1 Announce Type: new Abstract: This paper proposes a machine learning assisted portfolio optimization framework designed for low data environments and regime uncertainty. We construct

Predictions of charge density distributions for nuclei with Z geq 8

ResearchDGX agent

arXiv:2604.05312v1 Announce Type: cross Abstract: A deep neural network (DNN) has been developed to accurately predict nuclear charge density distributions for nuclei with proton numbers Z geq 8. By i

ProRe: A Proactive Reward System for GUI Agents via Reasoner-Actor Collaboration

SafetyDGX agent

arXiv:2509.21823v2 Announce Type: replace Abstract: Reward is critical to the evaluation and training of large language models (LLMs). However, existing rule-based or model-based reward methods strugg

ReasonScaffold: A Scaffolded Reasoning-based Annotation Protocol for Human-AI Co-Annotation

ResearchDGX agent

arXiv:2603.21094v3 Announce Type: replace Abstract: Human annotation is central to NLP evaluation, yet subjective tasks often exhibit substantial variability across annotators. While large language mo

Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards

Model ReleasesDGX agent

arXiv:2604.14876v1 Announce Type: cross Abstract: We study the tail behavior of regret in stochastic multi-armed bandits for algorithms that are asymptotically optimal in expectation. While minimizing

Revisiting Token Compression for Accelerating ViT-based Sparse Multi-View 3D Object Detectors

ResearchDGX agent

arXiv:2604.14563v1 Announce Type: new Abstract: Vision Transformer (ViT)-based sparse multi-view 3D object detectors have achieved remarkable accuracy but still suffer from high inference latency due

Right at My Level: A Unified Multilingual Framework for Proficiency-Aware Text Simplification

Model ReleasesDGX agent

arXiv:2604.05302v2 Announce Type: replace Abstract: Text simplification supports second language (L2) learning by providing comprehensible input, consistent with the Input Hypothesis. However, constru

SAQ: Stabilizer-Aware Quantum Error Correction Decoder

Model ReleasesDGX agent

arXiv:2512.08914v2 Announce Type: replace-cross Abstract: Quantum Error Correction (QEC) decoding faces a fundamental accuracy-efficiency tradeoff. Classical methods like Minimum Weight Perfect Matchi

Similarity-Distance-Magnitude Activations

ResearchDGX agent

arXiv:2509.12760v4 Announce Type: replace-cross Abstract: We introduce the Similarity-Distance-Magnitude (SDM) activation function, a more robust and interpretable formulation of the standard softmax

Social Story Frames: Contextual Reasoning about Narrative Intent and Reception

ResearchDGX agent

arXiv:2512.15925v2 Announce Type: replace Abstract: Reading stories evokes rich interpretive, affective, and evaluative responses, such as inferences about narrative intent or judgments about characte

The Cognitive Circuit Breaker: A Systems Engineering Framework for Intrinsic AI Reliability

ResearchDGX agent

arXiv:2604.13417v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed in mission-critical software systems, detecting hallucinations and ``faked truthfulness'' ha

The LLM Fallacy: Misattribution in AI-Assisted Cognitive Workflows

SafetyDGX agent

arXiv:2604.14807v1 Announce Type: cross Abstract: The rapid integration of large language models (LLMs) into everyday workflows has transformed how individuals perform cognitive tasks such as writing,

16 Apr 2026

A Bayesian Framework for Uncertainty-Aware Explanations in Power Quality Disturbance Classification

SafetyDGX agent

arXiv:2604.13658v1 Announce Type: new Abstract: Advanced deep learning methods have shown remarkable success in power quality disturbance (PQD) classification. To enhance model transparency, explainab

Activation-Guided Local Editing for Jailbreaking Attacks

SafetyDGX agent

arXiv:2508.00555v2 Announce Type: replace-cross Abstract: Jailbreaking is an essential adversarial technique for red-teaming these models to uncover and patch security flaws. However, existing jailbre

Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning

Model ReleasesDGX agent

arXiv:2512.08639v3 Announce Type: replace Abstract: Aerial Vision-and-Language Navigation (VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and navigate c

AgentSPEX: An Agent SPecification and EXecution Language

AgentsDGX agent

arXiv:2604.13346v1 Announce Type: new Abstract: Language-model agent systems commonly rely on reactive prompting, in which a single instruction guides the model through an open-ended sequence of reaso

Anima preview 3 , is it capable of uncensored images ? I am not getting any good results.

Local AiDGX agent

A Reddit thread on r/StableDiffusion discussing the capabilities and limitations of Anima Preview 3, a 2-billion-parameter text-to-image model by CircleStone Labs and Comfy Org, focused mainly on anim

ASTRA: Enhancing Multi-Subject Generation with Retrieval-Augmented Pose Guidance and Disentangled Position Embedding

Model ReleasesDGX agent

arXiv:2604.13938v1 Announce Type: new Abstract: Subject-driven image generation has shown great success in creating personalized content, but its capabilities are largely confined to single subjects i

Biased Federated Learning under Wireless Heterogeneity

SafetyDGX agent

arXiv:2503.06078v2 Announce Type: replace Abstract: Federated learning (FL) has emerged as a promising framework for distributed learning, enabling collaborative model training without sharing private

← Previous
1…494495496497498…1071
Next →