AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,952 results
22 Apr 2026

Introducing the Google Cloud Knowledge Catalog

Model ReleasesDGX agent

Traditional data catalogs were built as manual inventories for technical users, focusing on table structures rather than the deep context that AI agents need. When agents lack business semantics and d

Investigating Counterfactual Unfairness in LLMs towards Identities through Humor

SafetyDGX agent

arXiv:2604.18729v1 Announce Type: new Abstract: Humor holds up a mirror to social perception: what we find funny often reflects who we are and how we judge others. When language models engage with hum

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'her…

Model ReleasesDGX agent

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'hermes model' and select Kimi K2.6 to try out one of the most i

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LePREC: Reasoning as Classification over Structured Factors for Assessing Relevance of Legal Issues

Model ReleasesDGX agent

arXiv:2604.19464v1 Announce Type: cross Abstract: More than half of the global population struggles to meet their civil justice needs due to limited legal resources. While Large Language Models (LLMs)

Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2604.18897v1 Announce Type: new Abstract: We present a systematic empirical study of prompt engineering for formal mathematical reasoning in the context of the SAIR Equational Theories Stage 1 c

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part…

Model ReleasesDGX agent

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part is it doesn't use VLMs or any ML models at all. It's entire

LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

Model ReleasesDGX agent

arXiv:2604.19445v1 Announce Type: new Abstract: This paper presents a review for the LoViF Challenge on Real-World All-in-One Image Restoration. The challenge aimed to advance research on real-world a

MSDS: Deep Structural Similarity with Multiscale Representation

Model ReleasesDGX agent

arXiv:2604.19159v1 Announce Type: new Abstract: Deep-feature-based perceptual similarity models have demonstrated strong alignment with human visual perception in Image Quality Assessment (IQA). Howev

Multimodal Transformer for Sample-Aware Prediction of Metal-Organic Framework Properties

TutorialsDGX agent

arXiv:2604.19383v1 Announce Type: cross Abstract: Metal-organic frameworks (MOFs) are a major target of machine-learning-based property prediction, yet most models assume that a single framework repre

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

ApplicationsDGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

SafetyDGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

Quantum inspired qubit qutrit neural networks for real time financial forecasting

ResearchDGX agent

arXiv:2604.18838v1 Announce Type: new Abstract: This research investigates the performance and efficacy of machine learning models in stock prediction, comparing Artificial Neural Networks (ANNs), Qua

Reasoning Over Space: Enabling Geographic Reasoning for LLM-Based Generative Next POI Recommendation

Local AiDGX agent

arXiv:2601.04562v2 Announce Type: replace Abstract: Generative recommendation with large language models (LLMs) reframes prediction as sequence generation, yet existing LLM-based recommenders remain l

Reducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularization

ResearchDGX agent

arXiv:2604.19079v1 Announce Type: cross Abstract: Unification of automatic speech recognition (ASR) systems reduces development and maintenance costs, but training a single model to perform well in bo

SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution

Model ReleasesDGX agent

arXiv:2604.18982v1 Announce Type: new Abstract: Social intelligence, the ability to navigate complex interpersonal interactions, presents a fundamental challenge for language agents. Training such age

Streaming Structured Inference with Flash-SemiCRF

ResearchDGX agent

arXiv:2604.18780v1 Announce Type: new Abstract: Semi-Markov Conditional Random Fields (semi-CRFs) assign labels to segments of a sequence rather than to individual positions, enabling exact inference

TabReX : Tabular Referenceless eXplainable Evaluation

Model ReleasesDGX agent

arXiv:2512.15907v2 Announce Type: replace Abstract: Evaluating the quality of tables generated by large language models (LLMs) remains an open challenge: existing metrics either flatten tables into te

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

Model ReleasesDGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

Model ReleasesDGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

The new Gemini Enterprise: one platform for agent development, orchestration, and governance

Model ReleasesDGX agent

The first wave of AI changed how we find information; the next wave is changing how we get work done. Today, we’re enhancing our most powerful AI tools and bringing them together under one roof. Gemin

TrEEStealer: Stealing Decision Trees via Enclave Side Channels

ResearchDGX agent

arXiv:2604.18716v1 Announce Type: cross Abstract: Today, machine learning is widely applied in sensitive, security-related, and financially lucrative applications. Model extraction attacks undermine c

Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length

ResearchDGX agent

arXiv:2603.22608v2 Announce Type: replace Abstract: Users often rely on Large Language Models (LLMs) for processing multiple documents or performing analysis over a number of instances. For example, a

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

Model ReleasesDGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

VDPP: Video Depth Post-Processing for Speed and Scalability

Model ReleasesDGX agent

arXiv:2604.06665v2 Announce Type: replace Abstract: Video depth estimation is essential for providing 3D scene structure in applications ranging from autonomous driving to mixed reality. Current end-t

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

Model ReleasesDGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

What’s new in the Agentic Data Cloud: Powering the System of Action

Model ReleasesDGX agent

Companies are shifting from gen AI that simply answers questions to autonomous agents that perceive, reason, and act on their behalf. Attempting to scale these agents on legacy stacks exposes structur

When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification

Model ReleasesDGX agent

arXiv:2602.11199v2 Announce Type: replace Abstract: Large language models (LLMs) often respond even when prompts omit critical details or include misleading information, leading to hallucinations or r

21 Apr 2026

Access GPT Image 2.0 natively in Hermes Agent Update now to get access - just run `hermes update` and select your image generation tool mode…

AgentsDGX agent

Access GPT Image 2.0 natively in Hermes Agent Update now to get access - just run `hermes update` and select your image generation tool model with `hermes tools` Introducing ChatGPT Images 2.0 A state

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

Model ReleasesDGX agent

arXiv:2604.02846v2 Announce Type: replace Abstract: Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, c

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning

Model ReleasesDGX agent

arXiv:2604.17889v1 Announce Type: new Abstract: Despite recent progress in multimodal large language models (MLLMs), reliable visual question answering in aerial scenes remains challenging. In such sc

AeroScene: Progressive Scene Synthesis for Aerial Robotics

Model ReleasesDGX agent

arXiv:2603.23224v2 Announce Type: replace Abstract: Generative models have shown substantial impact across multiple domains, their potential for scene synthesis remains underexplored in robotics. This

ArgBench: Benchmarking LLMs on Computational Argumentation Tasks

Model ReleasesDGX agent

arXiv:2604.17366v1 Announce Type: new Abstract: Argumentation skills are an essential toolkit for large language models (LLMs). These skills are crucial in various use cases, including self-reflection

Benchmarking programs?

Local AiDGX agent

The Reddit post 'Benchmarking programs?' in r/ollama likely discusses tools and methods for measuring the performance of local language models running on Ollama. Available benchmarking tools for Ollam

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling

Local AiDGX agent

arXiv:2508.16745v2 Announce Type: replace Abstract: Reasoning is a core capability of large language models, yet how multi-step reasoning is learned and executed remains unclear. We study this questio

Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games

ResearchDGX agent

arXiv:2604.16785v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have enabled open-ended object recognition, yet they struggle with fine-grained tasks. In co

Calibrated? Not for Everyone: How Sexual Orientation and Religious Markers Distort LLM Accuracy and Confidence in Medical QA

ApplicationsDGX agent

arXiv:2604.17316v1 Announce Type: new Abstract: Safe clinical deployment of Large Language Models (LLMs) requires not only high accuracy but also robust uncertainty calibration to ensure models defer

Can we generate portable representations for clinical time series data using LLMs?

ApplicationsDGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

Model ReleasesDGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams

Model ReleasesDGX agent

arXiv:2604.16665v1 Announce Type: new Abstract: Urgent blood donation seeking posts and messages on social media often go unnoticed due to the overwhelming volume of daily communications. Traditional

Coevolving Representations in Joint Image-Feature Diffusion

ResearchDGX agent

arXiv:2604.17492v1 Announce Type: new Abstract: Joint image-feature generative modeling has recently emerged as an effective strategy for improving diffusion training by coupling low-level VAE latents

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

AgentsDGX agent

arXiv:2509.00789v2 Announce Type: replace Abstract: The pursuit of autonomous agents capable of temporally coherent planning is hindered by a fundamental flaw in current vision-language models (VLMs):

Compressing then Matching: An Efficient Pre-training Paradigm for Multimodal Embedding

ResearchDGX agent

arXiv:2511.08480v3 Announce Type: replace Abstract: Multimodal Large Language Models advance multimodal representation learning by acquiring transferable semantic embeddings, thereby substantially enh

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching

ResearchDGX agent

arXiv:2602.05449v3 Announce Type: replace Abstract: While diffusion models have achieved great success in the field of video generation, this progress is accompanied by a rapidly escalating computatio

DMax: Aggressive Parallel Decoding for dLLMs

SafetyDGX agent

arXiv:2604.08302v2 Announce Type: replace Abstract: We present DMax, a new paradigm for efficient diffusion language models (dLLMs). It mitigates error accumulation in parallel decoding, enabling aggr

Do LLM-derived graph priors improve multi-agent coordination?

Model ReleasesDGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

Document-as-Image Representations Fall Short for Scientific Retrieval

Model ReleasesDGX agent

arXiv:2604.18508v1 Announce Type: cross Abstract: Many recent document embedding models are trained on document-as-image representations, embedding rendered pages as images rather than the underlying

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

SafetyDGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

EgoSound: Benchmarking Sound Understanding in Egocentric Videos

Model ReleasesDGX agent

arXiv:2602.14122v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have recently achieved remarkable progress in vision-language understanding. Yet, human perception is inher

Enhancing Glass Surface Reconstruction via Depth Prior for Robot Navigation

Model ReleasesDGX agent

arXiv:2604.18336v1 Announce Type: cross Abstract: Indoor robot navigation is often compromised by glass surfaces, which severely corrupt depth sensor measurements. While foundation models like Depth A

EVE: Verifiable Self-Evolution of MLLMs via Executable Visual Transformations

SafetyDGX agent

arXiv:2604.18320v1 Announce Type: new Abstract: Self-evolution of multimodal large language models (MLLMs) remains a critical challenge: pseudo-label-based methods suffer from progressive quality degr

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

Model ReleasesDGX agent

arXiv:2508.07809v5 Announce Type: replace Abstract: Reinforcement learning with verifiable reward (RLVR) has become a promising paradigm for post-training large language models (LLMs) to improve their

FireScope: Wildfire Risk Prediction with a Chain-of-Thought Oracle

Model ReleasesDGX agent

arXiv:2511.17171v4 Announce Type: replace Abstract: Predicting wildfire risk is a reasoning-intensive spatial problem that requires the integration of visual, climatic, and geographic factors to infer

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready t…

Model ReleasesDGX agent

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready to fit the formats you need, from wide banners and presentati

Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens

ResearchDGX agent

arXiv:2604.17785v1 Announce Type: new Abstract: Unlearning in large language models (LLMs) has emerged as a promising safeguard against adversarial behaviors. When the forgetting loss is applied unifo

FRIGID: Scaling Diffusion-Based Molecular Generation from Mass Spectra at Training and Inference Time

Model ReleasesDGX agent

arXiv:2604.16648v1 Announce Type: new Abstract: In this work, we present FRIGID, a framework with a novel diffusion language model that generates molecular structures conditioned on mass spectra via i

From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization

ResearchDGX agent

arXiv:2601.16397v2 Announce Type: replace Abstract: Deploying multimodal large language models (MLLMs) for clinical summarization demands not only fluent generation but also transparency about where e

From log pi to pi: Taming Divergence in Soft Clipping via Bilateral Decoupled Decay of Probability Gradient Weight

Model ReleasesDGX agent

arXiv:2603.14389v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has catalyzed a leap in Large Language Model (LLM) reasoning, yet its optimization dynamics re

Geometric Stability: The Missing Axis of Representations

SafetyDGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

SafetyDGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

GR4CIL: Gap-compensated Routing for CLIP-based Class Incremental Learning

Model ReleasesDGX agent

arXiv:2604.17822v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) aims to continuously acquire new categories while preserving previously learned knowledge. Recently, Contrastive Langua

← Previous
1…365366367368369…1050
Next →