AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,098 results
10 Apr 2026

LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis

Model ReleasesDGX agent

arXiv:2510.24561v2 Announce Type: replace-cross Abstract: LoRA has become a widely adopted method for PEFT, and its initialization methods have attracted increasing attention. However, existing method

Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts?

Model ReleasesDGX agent

arXiv:2503.18018v2 Announce Type: replace Abstract: We demonstrate that large language models' (LLMs) mathematical reasoning is culturally sensitive: testing 14 models from Anthropic, OpenAI, Google,

LPM 1.0: Video-based Character Performance Model

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

Lumbermark: Resistant Clustering by Chopping Up Mutual Reachability Minimum Spanning Trees

Model ReleasesDGX agent

arXiv:2604.07143v1 Announce Type: new Abstract: We introduce Lumbermark, a robust divisive clustering algorithm capable of detecting clusters of varying sizes, densities, and shapes. Lumbermark iterat

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation

Model ReleasesDGX agent

arXiv:2604.06950v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are increasingly being deployed as automated content moderators. Within this landscape, we uncover a critic

Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS

Model ReleasesDGX agent

arXiv:2604.07276v1 Announce Type: cross Abstract: GROMACS is a de-facto standard for classical Molecular Dynamics (MD). The rise of AI-driven interatomic potentials that pursue near-quantum accuracy a

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

Model ReleasesDGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

Matrix Profile for Anomaly Detection on Multidimensional Time Series

Model ReleasesDGX agent

arXiv:2409.09298v2 Announce Type: replace-cross Abstract: The Matrix Profile (MP), a versatile tool for time series data mining, has been shown effective in time series anomaly detection (TSAD). This

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

Model ReleasesDGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts

Model ReleasesDGX agent

arXiv:2604.06505v1 Announce Type: cross Abstract: Large language models (LLMs) are widely explored for reasoning-intensive research tasks, yet resources for testing whether they can infer scientific c

MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

Model ReleasesDGX agent

arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

Model ReleasesDGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

MinerU2.5-Pro: Pushing the Limits of Data-Centric Document Parsing at Scale

Model ReleasesDGX agent

arXiv:2604.04771v2 Announce Type: replace-cross Abstract: Current document parsing methods advance primarily through model architecture innovation, while systematic engineering of training data remain

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

Model ReleasesDGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

Mitigating Spurious Background Bias in Multimedia Recognition with Disentangled Concept Bottlenecks

Model ReleasesDGX agent

arXiv:2510.15770v3 Announce Type: replace Abstract: Concept Bottleneck Models (CBMs) enhance interpretability by predicting human-understandable concepts as intermediate representations. However, exis

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

Model ReleasesDGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

Mixed-Initiative Context: Structuring and Managing Context for Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2604.07121v1 Announce Type: cross Abstract: In the human-AI collaboration area, the context formed naturally through multi-turn interactions is typically flattened into a chronological sequence

mlx @ollama!

Model ReleasesDGX agent

Ollama released a preview version (0.19) on March 31, 2026, built on top of Apple's open-source MLX framework, enabling local LLMs to run significantly faster on Apple Silicon Macs by leveraging th...

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation

Model ReleasesDGX agent

arXiv:2601.02535v2 Announce Type: replace Abstract: Selecting a single high-quality output from multiple stochastic generations remains a fundamental challenge for large language models (LLMs), partic

Monocular Depth Estimation From the Perspective of Feature Restoration: A Diffusion Enhanced Depth Restoration Approach

Model ReleasesDGX agent

arXiv:2604.07664v1 Announce Type: new Abstract: Monocular Depth Estimation (MDE) is a fundamental computer vision task with important applications in 3D vision. The current mainstream MDE methods empl

Multi-Agent Orchestration is a great tool in building agentic products thankfully @sydneyrunkle put together a guide with runnable code for …

Model ReleasesDGX agent

Multi-Agent Orchestration is a great tool in building agentic products thankfully @sydneyrunkle put together a guide with runnable code for each pattern earlier this year 🙏 (link below) choose any mod

Multi-objective Evolutionary Merging Enables Efficient Reasoning Models

Model ReleasesDGX agent

arXiv:2604.06465v1 Announce Type: cross Abstract: Reasoning models have demonstrated remarkable capabilities in solving complex problems by leveraging long chains of thought. However, this more delibe

Near-100% Accurate Data for your Agent with Comprehensive Context Engineering

Model ReleasesDGX agent

Agentic workflows are already used for initiating action. To be successful, agents typically need to combine multiple steps and execute business logic reflective of real-life decisions. But, as develo

Negative Binomial Variational Autoencoders for Overdispersed Latent Modeling

Model ReleasesDGX agent

arXiv:2508.05423v2 Announce Type: replace Abstract: Although artificial neural networks are often described as brain-inspired, their representations typically rely on continuous activations, such as t

Neural parametric representations for thin-shell shape optimisation

Model ReleasesDGX agent

arXiv:2604.06612v1 Announce Type: cross Abstract: Shape optimisation of thin-shell structures requires a flexible, differentiable geometric representation suitable for gradient-based optimisation. We

New in Claude Code: /ultraplan Claude builds an implementation plan for you on the web. You can read it and edit it, then run the plan on th…

Model ReleasesDGX agent

New in Claude Code: /ultraplan Claude builds an implementation plan for you on the web. You can read it and edit it, then run the plan on the web or back in your terminal. Available now in preview for

nice AIE weekend bonus: @barry_zyj and @MaheshMurag have officially become the second AIE Millionaire speakers!! they reached this milestone…

Model ReleasesDGX agent

nice AIE weekend bonus: @barry_zyj and @MaheshMurag have officially become the second AIE Millionaire speakers!! they reached this milestone 2x as fast as @sgrove, our first AIE Millionaire... ... may

No, Mythos is not *that* good; its PR is that good. 🙄

Model ReleasesDGX agent

No, Mythos is not *that* good; its PR is that good. 🙄 *BESSENT SUMMONED WALL STREET CEOS TO DISCUSS ANTHROPIC’S MYTHOS This is crazy. Claude Mythos is apparently so good that Bessent and J. Powell sum

Noise Immunity in In-Context Tabular Learning: An Empirical Robustness Analysis of TabPFN's Attention Mechanisms

Model ReleasesDGX agent

arXiv:2604.04868v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs) such as TabPFN (Tabular Prior-Data Fitted Network) are designed to generalize across heterogeneous tabular da

Not All Tokens See Equally: Perception-Grounded Policy Optimization for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.01840v2 Announce Type: replace Abstract: While Reinforcement Learning from Verifiable Rewards (RLVR) has advanced reasoning in Large Vision-Language Models (LVLMs), prevailing frameworks su

NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration: Methods and Results

Model ReleasesDGX agent

arXiv:2604.06945v2 Announce Type: replace Abstract: This paper reports on the NTIRE 2026 Challenge on Bitstream-Corrupted Video Restoration (BSCVR). The challenge aims to advance research on recoverin

ODYN: An All-Shifted Non-Interior-Point Method for Quadratic Programming in Robotics and AI

Model ReleasesDGX agent

arXiv:2602.16005v2 Announce Type: replace-cross Abstract: We introduce ODYN, a novel all-shifted primal-dual non-interior-point quadratic programming (QP) solver designed to efficiently handle challen

OmniTabBench: Mapping the Empirical Frontiers of GBDTs, Neural Networks, and Foundation Models for Tabular Data at Scale

Model ReleasesDGX agent

arXiv:2604.06814v1 Announce Type: cross Abstract: While traditional tree-based ensemble methods have long dominated tabular tasks, deep neural networks and emerging foundation models have challenged t

On Emotion-Sensitive Decision Making of Small Language Model Agents

Model ReleasesDGX agent

arXiv:2604.06562v1 Announce Type: new Abstract: Small language models (SLM) are increasingly used as interactive decision-making agents, yet most decision-oriented evaluations ignore emotion as a caus

On Integrating Resilience and Human Oversight into LLM-Assisted Modeling Workflows for Digital Twins

Model ReleasesDGX agent

arXiv:2603.25898v2 Announce Type: replace-cross Abstract: LLM-assisted modeling holds the potential to rapidly build executable Digital Twins of complex systems from only coarse descriptions and senso

On-Policy Distillation of Language Models for Autonomous Vehicle Motion Planning

Model ReleasesDGX agent

arXiv:2604.07944v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong potential for autonomous vehicle motion planning by reformulating trajectory prediction a

Open-Ended Instruction Realization with LLM-Enabled Multi-Planner Scheduling in Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2604.08031v1 Announce Type: cross Abstract: Most Human-Machine Interaction (HMI) research overlooks the maneuvering needs of passengers in autonomous driving (AD). Natural language offers an int

Optimal Rates for Pure {arepsilon}-Differentially Private Stochastic Convex Optimization with Heavy Tails

Model ReleasesDGX agent

arXiv:2604.06492v1 Announce Type: new Abstract: We study stochastic convex optimization (SCO) with heavy-tailed gradients under pure epsilon-differential privacy (DP). Instead of assuming a bound on t

Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models

Model ReleasesDGX agent

arXiv:2604.08266v1 Announce Type: new Abstract: Leveraging the general world knowledge of Large Language Models (LLMs) holds significant promise for improving the ability of autonomous driving systems

oslash Source Models Leak What They Shouldn't nrightarrow: Unlearning Zero-Shot Transfer in Domain Adaptation Through Adversarial Optimization

Model ReleasesDGX agent

arXiv:2604.08238v1 Announce Type: new Abstract: The increasing adaptation of vision models across domains, such as satellite imagery and medical scans, has raised an emerging privacy risk: models may

Our first successful Gemma 4 Runtime in London with @swyx @patloeber @nick_kango @cormacb and others! 💎Great to go out for a run and talk a…

Model ReleasesDGX agent

Our first successful Gemma 4 Runtime in London with @swyx @patloeber @nick_kango @cormacb and others! 💎Great to go out for a run and talk about Gemma, agents, evals and more @osanseviero @swyx and oth

Our response to the Axios developer tool compromise

Model ReleasesDGX agent

Axios, a widely used third-party JavaScript developer library with approximately 100 million weekly downloads, was compromised on March 31, 2026, as part of a broader software supply chain attack a...

ParseBench: A Document Parsing Benchmark for AI Agents

Model ReleasesDGX agent

arXiv:2604.08538v1 Announce Type: new Abstract: AI agents are changing the requirements for document parsing. What matters is semantic correctness: parsed output must preserve the structure and

PASK: Toward Intent-Aware Proactive Agents with Long-Term Memory

Model ReleasesDGX agent

arXiv:2604.08000v1 Announce Type: cross Abstract: Proactivity is a core expectation for AGI. Prior work remains largely confined to laboratory settings, leaving a clear gap in real-world proactive age

People ask me how I choose what model to route queries to. It's simple. Claude gets knowledge work anything below that would be degrading Cl…

Model ReleasesDGX agent

I was unable to retrieve the specific tweet content from that URL, as the post requires a logged-in X (Twitter) session to access, and search results did not surface the full text of that specific ...

PeReGrINE: Evaluating Personalized Review Fidelity with User Item Graph Context

Model ReleasesDGX agent

arXiv:2604.07788v1 Announce Type: cross Abstract: We introduce PeReGrINE, a benchmark and evaluation framework for personalized review generation grounded in graph-structured user--item evidence. PeRe

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

Model ReleasesDGX agent

arXiv:2604.07343v1 Announce Type: cross Abstract: Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a cen

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2512.23994v3 Announce Type: replace-cross Abstract: Text-to-audio-video (T2AV) generation is central to applications such as filmmaking and world modeling. However, current models often fail to

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

Model ReleasesDGX agent

arXiv:2604.07230v2 Announce Type: replace Abstract: Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However,

Physical Knot Classification Beyond Accuracy: A Benchmark and Diagnostic Study

Model ReleasesDGX agent

arXiv:2603.23286v3 Announce Type: replace Abstract: Physical knot classification is a fine-grained task in which the intended cue is rope crossing structure, but high accuracy may still come from appe

Physics-Informed Functional Link Constrained Framework with Domain Mapping for Solving Bending Analysis of an Exponentially Loaded Perforated Beam

Model ReleasesDGX agent

arXiv:2604.07025v1 Announce Type: cross Abstract: This article presents a novel and comprehensive approach for analyzing bending behavior of the tapered perforated beam under an exponential load. The

Physics-Informed Neural Networks for Joint Source and Parameter Estimation in Advection-Diffusion Equations

Model ReleasesDGX agent

arXiv:2512.07755v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated the success of deep learning in solving forward and inverse problems in engineering and scientific computing

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

Model ReleasesDGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

Pinecone now has a native extension for Gemini CLI. gemini extensions install https://github.com/pinecone-io/gemini-cli-extension → 9 MCP to…

Model ReleasesDGX agent

Pinecone now has a native extension for Gemini CLI. gemini extensions install https://github.com/pinecone-io/gemini-cli-extension → 9 MCP tools → 7 skills (quickstart, query, assistant, RAG and more)

Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks

Model ReleasesDGX agent

arXiv:2511.19474v4 Announce Type: replace-cross Abstract: Automatically detecting abnormal events in videos is crucial for modern autonomous systems, yet existing Video Anomaly Detection (VAD) benchma

Plasma GraphRAG: Physics-Grounded Parameter Selection for Gyrokinetic Simulations

Model ReleasesDGX agent

arXiv:2604.06279v1 Announce Type: cross Abstract: Accurate parameter selection is fundamental to gyrokinetic plasma simulations, yet current practices rely heavily on manual literature reviews, leadin

PLUME: Latent Reasoning Based Universal Multimodal Embedding

Model ReleasesDGX agent

arXiv:2604.02073v2 Announce Type: replace Abstract: Universal multimodal embedding (UME) maps heterogeneous inputs into a shared retrieval space with a single model. Recent approaches improve UME by g

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.08340v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved remarkable progress in static visual understanding, their deployment in complex 3D embodied environmen

Possible memory leak in Ollama when using Claude Code?

Model ReleasesDGX agent

Users in the r/ollama community have reported a possible memory leak occurring in Ollama when it is used as a backend with Claude Code, with Ollama runner processes not always being properly termin...

← Previous
1…362363364365366…369
Next →