AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

Parameter Golf: What Really Works?

DGX agent

arXiv:2607.01517v1 Announce Type: new Abstract: How far can a language model improve under a strict artifact budget? Parameter Golf posed this question as an open community challenge in which particip

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.02426v1 Announce Type: cross Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensi

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Spectral Imbalance Causes Forgetting in Low-Rank Continual Adaptation

DGX agent

arXiv:2602.00722v2 Announce Type: replace Abstract: Parameter-efficient continual learning aims to adapt pre-trained models to sequential tasks without forgetting previously acquired knowledge. Most e

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Text-Driven 3D Indoor Scene Synthesis in Non-Manhattan Environments

DGX agent

arXiv:2607.02407v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in 3D indoor synthesis for Manhattan environments. However, existing methods ofte

model-releasesarxiv-cs-ai
3 Jul 2026
Research

Will Scaling Improve Social Simulation with LLMs?

DGX agent

arXiv:2607.02464v1 Announce Type: new Abstract: Large Language Model (LLM) social simulations are a promising research method, but they are not yet faithful enough to be adopted widely. In this work,

researcharxiv-cs-cl
3 Jul 2026
Model Releases

AGC-Bench: Measuring Artificial General Creativity

DGX agent

arXiv:2607.01152v1 Announce Type: new Abstract: Creativity research has debated whether creativity is domain-specific (e.g., visual, writing, science), and if it is psychometrically separable from gen

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

AlgoBench: Benchmarking Algorithmic Adaptation in Code Generation

DGX agent

arXiv:2607.00062v1 Announce Type: cross Abstract: High pass rates on established programming benchmarks such as HumanEval and LiveCodeBench do not always show whether a model can reason about algorith

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Beyond Activation Alignment:The Alignment-Diversity Tradeoff in Task-Aware LLM Quantization

DGX agent

arXiv:2607.00908v1 Announce Type: new Abstract: Mixed-precision quantization (MPQ) has become a key technique for deploying large language models under stringent memory and compute constraints. We fir

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Clinician-Level Agreement Without Clinical Caution: LLM Evaluator Limits in Medical AI Benchmarking

DGX agent

arXiv:2607.01103v1 Announce Type: new Abstract: Open-response evaluation provides stronger clinical validity than multiple-choice benchmarks but creates a scoring bottleneck that motivates automated L

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images

DGX agent

arXiv:2607.00338v1 Announce Type: new Abstract: Object detection for Unmanned Aerial Vehicles (UAVs) working in open and dynamic environments is a highly challenging task. While Vision-Language Models

model-releasesarxiv-cs-cv
2 Jul 2026
Research

Efficient Multilingual Reasoning Transfer via Progressive Code-Switching

DGX agent

arXiv:2607.00485v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong reasoning capabilities in English, yet their performance degrades significantly when required to reas

researcharxiv-cs-cl
2 Jul 2026
Model Releases

GeoSearcher: Anchor-Guided Progressive Reasoning for Remote Sensing Visual Grounding with Process Supervision

DGX agent

arXiv:2607.01050v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown strong cross-modal understanding and coordinate generation abilities in visual grounding. How

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Ghost in the Kernel: In-Context Learning with Efficient Transformers via Domain Generalization

DGX agent

arXiv:2607.00479v1 Announce Type: new Abstract: Transformer-based large models have demonstrated remarkable generalization abilities across different tasks by leveraging a context-aware attention modu

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

GRACE-RAG: Governed Retrieval Architecture for Canonical Evidence Synthesis, Enabling Lightweight Deployment in Closed-Domain Institutional Settings

DGX agent

arXiv:2607.00013v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are widely used in institutional question answering settings where responses must be grounded in authorit

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

GryphOne: Symbol-Aware Masked Diffusion for Structural Refinement in Offline Handwritten Mathematical Expression Recognition

DGX agent

arXiv:2602.03370v2 Announce Type: replace Abstract: Handwritten mathematical expression recognition (HMER) requires reasoning over diverse symbols and structures, yet autoregressive models struggle wi

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Identifying and Resolving Pitfalls of Knowledge-Based VQA Benchmarks: Auditing, Repairing, and Augmenting

DGX agent

arXiv:2607.00159v1 Announce Type: new Abstract: Knowledge-Based Visual Question Answering (KB-VQA) aims to evaluate whether Visual Language Models (VLMs) can retrieve, ground, and reason over external

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Information-Regularized Attention for Visual-Centric Reasoning

DGX agent

arXiv:2607.00434v1 Announce Type: new Abstract: Vision-language models (VLMs) have become a paradigm for multimodal learning, yet remain unstable due to object hallucination, weak visual grounding, an

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

MolSafeEval: A Benchmark for Uncovering Safety Risks in AI-Generated Molecules

DGX agent

arXiv:2607.00464v1 Announce Type: cross Abstract: Current molecular generation benchmarks emphasize task complexity, molecule novelty, and property alignment; they largely overlook a critical concern:

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

MSQA: A Natively Sourced Multilingual and Multicultural SimpleQA Benchmark

DGX agent

arXiv:2607.00724v1 Announce Type: new Abstract: Multilingual fluency often invites a stronger assumption: a model that can speak a user's language must also understand the culture encoded by that lang

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

Not All Prediction Targets Keep Training-Free Diffusion Guidance on the Manifold

DGX agent

arXiv:2607.00647v1 Announce Type: new Abstract: Training-free guidance (TFG) steers a pretrained diffusion model toward a desired attribute at inference. To be effective, this guidance must be applied

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

On the Reliability of Cue Conflict and Beyond

DGX agent

arXiv:2603.10834v4 Announce Type: replace-cross Abstract: Understanding how neural networks rely on visual cues offers a human-interpretable view of their internal decision processes. The cue-conflict

model-releasesarxiv-cs-ai
2 Jul 2026
Research

OpFML: Pipeline for ML-based Operational Inference

DGX agent

arXiv:2601.11046v2 Announce Type: replace Abstract: Machine learning models for climate and Earth science are becoming increasingly capable, yet model deployment into operational use remains a largely

researcharxiv-cs-lg
2 Jul 2026
Model Releases

PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking

DGX agent

arXiv:2607.00115v1 Announce Type: new Abstract: This paper explores multi-turn visual reasoning and observes that MLLMs repeatedly fail to localize the target, leading to long, redundant trajectories.

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Quantifying the Affective Gap: A Zero-Shot Evaluation of LLMs on Fine-Grained Emotion Taxonomies

DGX agent

arXiv:2607.00968v1 Announce Type: new Abstract: Emotion recognition in natural language is a foundational challenge in affective computing, with critical implications for human-computer interaction, m

model-releasesarxiv-cs-cl
2 Jul 2026
Safety

RoboWorld: Fast and Reliable Neural Simulators for Generalist Robot Policy Evaluation

DGX agent

arXiv:2607.01060v1 Announce Type: new Abstract: Video world models are emerging as a scalable alternative for evaluating generalist robot policies, bypassing the physical constraints and engineering b

safetyarxiv-cs-ro
2 Jul 2026
Model Releases

Skills Are Not Islands: Measuring Dependency and Risk in Agent Skill Supply Chains

DGX agent

arXiv:2607.01136v1 Announce Type: cross Abstract: Agent skills package reusable operational knowledge for Large Language Model (LLM) agents, yet as they grow in scope, they become dependency-bearing a

model-releasesarxiv-cs-ai
2 Jul 2026
Research

Slope-Guided Mamba and Angular-Refined Transformer for Light Field Super-Resolution

DGX agent

arXiv:2607.00965v1 Announce Type: new Abstract: Light Field Super-Resolution (LFSR) necessitates accurate modeling of spatial-angular correlations while preserving intrinsic 4D ray coherence. However,

researcharxiv-cs-cv
2 Jul 2026
Model Releases

StochasT: Learning with Stochastic Turn Depth for Visual Instruction Tuning

DGX agent

arXiv:2607.00465v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely extensively on Visual Instruction Tuning (VIT) to elicit their multimodal reasoning capabilities. However, w

model-releasesarxiv-cs-cl
2 Jul 2026
Local Ai

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks

DGX agent

arXiv:2607.00053v1 Announce Type: cross Abstract: Large language models (LLMs) embedded in multi-turn agentic harnesses are reshaping software engineering (SWE), but routing every task to a frontier m

local-aiarxiv-cs-ai
2 Jul 2026
Model Releases

Tail-Shape Estimation in LLM Evaluation Is Fragile: A Protocol for Diagnosing False Positives

DGX agent

arXiv:2606.16511v2 Announce Type: replace Abstract: Recent work motivates moving large language model (LLM) evaluation from mean-based to tail-aware metrics, including conditional value-at-risk and ta

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

TallyTrain: Communication-Efficient Federated Distillation

DGX agent

arXiv:2607.00173v1 Announce Type: new Abstract: Federated learning is bandwidth-bound on two orthogonal axes: model size, which limits how often parameter-averaging methods can afford to merge, and cl

model-releasesarxiv-cs-lg
2 Jul 2026
Research

Towards Accurate State Estimation: Motion Dynamics Kalman Filter for 3D Multi-Object Tracking

DGX agent

arXiv:2505.07254v2 Announce Type: replace Abstract: Precise 3D state estimation in multi-object tracking (MOT) is critical for self-driving cars, particularly for objects occluded. Motion modeling in

researcharxiv-cs-cv
2 Jul 2026
Model Releases

TRIE: An Evaluation Framework for Stochastic PDE Surrogates

DGX agent

arXiv:2607.00196v1 Announce Type: new Abstract: Many scientific systems exhibit uncertainty from stochastic forcing, unresolved degrees of freedom, or imperfect observations, making reliable surrogate

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Understanding How Humans Inject Knowledge into Machine Learning Workflows through Visual Analytics

DGX agent

arXiv:2607.00969v1 Announce Type: cross Abstract: Visual analytics (VA) plays an increasingly important role in supporting machine learning (ML) workflows. In the field of visualization, such approach

model-releasesarxiv-cs-lg
2 Jul 2026
Research

Vitality-Aware Compression for Efficient Image-to-Shape Diffusion Transformers

DGX agent

arXiv:2607.00382v1 Announce Type: new Abstract: We propose the first compression approach for image-to-shape Diffusion Transformers (DiTs) that substantially reduces model size while preserving geomet

researcharxiv-cs-cv
2 Jul 2026
Model Releases

ZO-Act: Efficient Zeroth-Order Fine-Tuning via One-Shot Activation-Informed Low-Rank Subspaces

DGX agent

arXiv:2607.01125v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization enables fine-tuning large language models when backpropagation is unavailable or memory-prohibitive, but existing methods

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

A Systematic Approach to Multi-Agent AI from Advanced Regulatory Control Theory: Safe and Auditable LLM Operator Agents for Process Control

DGX agent

arXiv:2606.30877v1 Announce Type: cross Abstract: Recent literature shows that large language models (LLMs) are useful for general-purpose tasks yet perform poorly on specific domain ones. One reason

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Beyond Clean Text: Evaluating Encoder and Decoder Robustness for Bangla Event Detection in Noisy Text

DGX agent

arXiv:2606.30914v1 Announce Type: new Abstract: Event detection (ED) systems are typically evaluated on clean, curated text, leaving their robustness to real-world noise largely unexplored, particular

model-releasesarxiv-cs-cl
1 Jul 2026
Local Ai

ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Industries

DGX agent

arXiv:2606.31163v1 Announce Type: cross Abstract: Large language models deployed in regulated industries operate under two constraints: compliance enforcement and cost efficiency. Personally identifia

local-aiarxiv-cs-ai
1 Jul 2026
Model Releases

DynFly: Dynamic-Aware Continuous Trajectory Generation for UAV Vision-Language Navigation in Urban Environments

DGX agent

arXiv:2606.31654v1 Announce Type: cross Abstract: Recent advances in multimodal large models have significantly improved UAV vision-language navigation (UAV-VLN) by enhancing high-level perception and

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Energy-Optimal Spatial Iterative Learning within a Virtual Tube

DGX agent

arXiv:2606.31487v1 Announce Type: new Abstract: Due to the limited endurance of embedded energy sources such as lithium-polymer (LiPo) batteries, the flight duration and operational range of unmanned

model-releasesarxiv-cs-ro
1 Jul 2026
Model Releases

Evidence Triangulation for Multimodal Fact-Checking in the Wild

DGX agent

arXiv:2606.31367v1 Announce Type: cross Abstract: The proliferation of multimedia content on social platforms has fueled multimodal misinformation, where images are used to reinforce false claims. Con

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

Evo-PI: Aligning Medical Reasoning via Evolving Principle-Guided Supervision

DGX agent

arXiv:2606.31800v1 Announce Type: new Abstract: Despite recent progress, the reasoning capabilities of large multimodal language models (MLLMs) remain fundamentally constrained by static supervision,

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

FinPersona-Bench: A Benchmark for Longitudinal Psychometric Stability of Autonomous Financial Agents

DGX agent

arXiv:2606.31522v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous financial agents initialized with explicit behavioral mandates such as 'preserve

model-releasesarxiv-cs-ai
1 Jul 2026
Local Ai

Histogram-constrained Image Generation

DGX agent

arXiv:2606.31683v1 Announce Type: cross Abstract: Diffusion models have emerged as a dominant paradigm in generative modeling, enabling high-fidelity sampling from complex data distributions. Despite

local-aiarxiv-cs-ai
1 Jul 2026
Model Releases

M3CoTBench: Benchmark Chain-of-Thought of MLLMs in Medical Image Understanding

DGX agent

arXiv:2601.08758v4 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) reasoning has proven effective in enhancing large language models by encouraging step-by-step intermediate reasoning, a

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

DGX agent

arXiv:2606.31966v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have strong potential as embodied agents, but their ability to collaborate in visually grounded enviro

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Medical Image Spatial Grounding with Semantic Sampling

DGX agent

arXiv:2603.14579v3 Announce Type: replace Abstract: Vision language models (VLMs) have shown significant promise in visual grounding for images as well as videos. In medical imaging research, VLMs rep

model-releasesarxiv-cs-cv
1 Jul 2026
← Previous
1…349350351352353…1067
Next →