AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

IFAR: Multi-Perspective and Multi-Level Causal Discovery with LLMs

DGX agent

arXiv:2409.05559v2 Announce Type: replace Abstract: Large language models (LLMs) have developed rapidly, and their reasoning capabilities have become a hot research topic. However, there is still limi

model-releasesarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

DGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

DGX agent

arXiv:2607.08540v1 Announce Type: cross Abstract: Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to to

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Infinity-Parser2 Technical Report

DGX agent

arXiv:2607.07836v1 Announce Type: new Abstract: We present Infinity-Parser2, a large multimodal model that couples a controllable data-synthesis pipeline with multi-task reinforcement learning for end

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE

DGX agent

arXiv:2607.07740v1 Announce Type: cross Abstract: Modern LLMs are increasingly deployed in long-context applications such as retrieval-augmented generation, repository-level coding, and agentic workfl

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Joint Bayesian Parameter and Model Order Estimation for Low-Rank Probability Mass Tensors

DGX agent

arXiv:2410.06329v4 Announce Type: replace-cross Abstract: Obtaining a reliable estimate of the joint probability mass function (PMF) of a set of random variables from observed data is a significant ob

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Joint Discrete-Continuous Flow Matching for Open-Vocabulary Inverse Design of Multilayer Optical Coatings

DGX agent

arXiv:2607.08392v1 Announce Type: cross Abstract: Amortized neural inverse design typically remains closed-world: component choices are fixed vocabulary tokens, coordinate grids are frozen at training

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

KronQ: LLM Quantization via Kronecker-Factored Hessian

DGX agent

arXiv:2607.07964v1 Announce Type: new Abstract: Post-training quantization (PTQ) is a widely adopted technique for compressing large language models (LLMs) without retraining. Existing second-order PT

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning

DGX agent

arXiv:2603.12478v2 Announce Type: replace Abstract: Multimodal instruction tuning is often compute-inefficient because training budgets are spread across large mixed image-video pools whose utility is

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity

DGX agent

arXiv:2607.08152v1 Announce Type: cross Abstract: On the recent EyeBench benchmark, predicting reading comprehension from eye movements exposes a stark gap: text-aware models using pretrained language

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

LightCrafter: PBR-Conditioned Video Diffusion Refinement for Controllable and Consistent Relighting

DGX agent

arXiv:2607.08016v1 Announce Type: new Abstract: Video relighting requires balancing long-form temporal consistency with a physically grounded understanding of light transport, which depends on accurat

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing

DGX agent

arXiv:2607.07953v1 Announce Type: cross Abstract: Self-attention lets each token retrieve information from the full context, but its quadratic cost in sequence length limits training and inference at

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

LiST: Lipschitz Scaling Training for Robust and Calibrated Neural Networks

DGX agent

arXiv:2607.07745v1 Announce Type: new Abstract: While accuracy, robustness, and calibration are all essential for reliable neural networks, they are often studied separately; developing models that sa

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

LlamaSeg: Image Segmentation via Autoregressive Mask Generation

DGX agent

arXiv:2505.19422v2 Announce Type: replace Abstract: We present extbf{LlamaSeg}, a visual autoregressive framework that unifies multiple image segmentation tasks via natural language instructions. By r

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

LSRM: High-Fidelity Object-Centric Reconstruction via Scaled Context Windows

DGX agent

arXiv:2604.05182v2 Announce Type: replace-cross Abstract: We introduce the Large Sparse Reconstruction Model to study how scaling transformer context windows affects feed-forward 3D reconstruction. Al

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

LUMI: Tokenizer-Agnostic LLM-Based Lossless Image Compression

DGX agent

arXiv:2607.08221v1 Announce Type: new Abstract: Large language model (LLM)-based lossless image compression methods typically represent pixel data through the native text interface of a pretrained mod

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

MetaHGNIE: Meta-Path Induced Hypergraph Contrastive Learning in Heterogeneous Knowledge Graphs

DGX agent

arXiv:2512.12477v2 Announce Type: replace Abstract: Estimating node importance in heterogeneous knowledge graphs is a fundamental problem underlying recommendation, search, and knowledge decision syst

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Mixture of Enhanced-View Experts for Multi-Query Vehicle ReID and A Large-Scale Benchmark

DGX agent

arXiv:2607.08085v1 Announce Type: new Abstract: Multi-query vehicle ReID aims to leverage complementary information from diverse views for robust feature learning. However, current methods suffer from

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

MSRNet: A Multi-Scale Recursive Network for Camouflaged Object Detection

DGX agent

arXiv:2511.12810v2 Announce Type: replace-cross Abstract: Camouflaged object detection is an emerging and challenging computer vision task that requires identifying and segmenting objects that blend s

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Multi-Resolution Feature Stem for Diabetic Retinopathy lesion segmentation

DGX agent

arXiv:2607.08679v1 Announce Type: new Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness worldwide, requiring automated lesion segmentation using deep learning models for

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

DGX agent

arXiv:2607.08423v1 Announce Type: new Abstract: The rapid integration of Large Vision-Language Models (VLMs) into critical infrastructure promises to revolutionize personalized healthcare and dietary

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

On the Role of Conversational Timing in Synthetic Training Data for ASR

DGX agent

arXiv:2607.08371v1 Announce Type: cross Abstract: Synthetic multi-speaker conversations are widely used to train conversational automatic speech recognition (ASR) systems, but it remains unclear which

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

ParamMute: Suppressing Knowledge-Critical FFNs for Faithful Retrieval-Augmented Generation

DGX agent

arXiv:2502.15543v4 Announce Type: replace-cross Abstract: Large language models (LLMs) integrated with retrieval-augmented generation (RAG) have improved factuality by grounding outputs in external ev

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

path_boost: A Python Package for Interpretable Graph-Level Prediction using Path-Based Gradient Boosting

DGX agent

arXiv:2607.07935v1 Announce Type: cross Abstract: We present path_boost, a Python package for interpretable supervised learning on graph-structured input data. The package implements PathBoost, a grad

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Persistent Multiscale Density-based Clustering

DGX agent

arXiv:2512.16558v3 Announce Type: replace Abstract: Clustering is a cornerstone of modern data analysis. Detecting clusters in exploratory data analyses (EDA) requires algorithms that make few assumpt

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring

DGX agent

arXiv:2607.08066v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is a promising safety mechanism for AI agents, based on the premise that visible reasoning traces can surface misalign

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

PhasorFlow: A Python Library for Unit Circle Based Computing

DGX agent

arXiv:2603.15886v3 Announce Type: replace-cross Abstract: We present PhasorFlow, an open-source Python library for computing on the S^1 unit circle. Inputs are encoded as complex phasors z=e^{iphi} on

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Physics-Informed Machine Learning Under Small-Data Constraints: Lessons from Abrasive Waterjet Milling

DGX agent

arXiv:2607.07863v1 Announce Type: new Abstract: In physically dominated machining processes, experimental datasets are small, expensive, and material-specific; in this regime, data curation, evaluatio

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

PolyUQuest: Verifiable Structure-Aware Web RAG over Heterogeneous Graphs

DGX agent

arXiv:2607.08269v1 Announce Type: new Abstract: Existing retrieval-augmented generation (RAG) systems treat web pages as flat text, losing the structural and semantic signals encoded in HTML. We prese

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Pose-to-Biomechanics: Bridging 3D Human Pose Estimation and Biomechanical Attribute Prediction

DGX agent

arXiv:2607.08725v1 Announce Type: cross Abstract: Recent progress in 3D human pose estimation has made markerless recovery of skeletal motion increasingly accurate and scalable. However, most pose est

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Psychological Competence as a Missing Dimension in AI Evaluation

DGX agent

arXiv:2607.08285v1 Announce Type: new Abstract: Current AI evaluation frameworks focus primarily on technical performance, including accuracy, robustness, reasoning ability, and policy compliance. The

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Real-World Blind Super-Resolution via Feature Matching with Implicit High-Resolution Priors

DGX agent

arXiv:2202.13142v3 Announce Type: replace Abstract: A key challenge of real-world image super-resolution (SR) is to recover the missing details in low-resolution (LR) images with complex unknown degra

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

ReCoLoRA: Spectrum-Aware Recursive Consolidation for Continual LLM Fine-Tuning

DGX agent

arXiv:2607.07719v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning adapts a large language model to one task cheaply, but across a task sequence LoRA-style methods keep stacking low-ran

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Reinforcing the Generation Order of Multimodal Masked Diffusion Models

DGX agent

arXiv:2607.08056v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) have recently achieved substantial progress in natural language generation tasks. Recent research demonstrates that a

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents

DGX agent

arXiv:2607.08716v1 Announce Type: new Abstract: In long-horizon tasks, decision-relevant state is often scattered across an expanding trajectory, while the action agent must surface it and act. As tra

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models

DGX agent

arXiv:2607.08665v1 Announce Type: new Abstract: Routing among large language models (LLMs) trades response quality against serving cost, motivated by the reported gap between deployed routers and a pe

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments

DGX agent

arXiv:2603.16453v3 Announce Type: replace Abstract: Large language model (LLM) agents have made rapid progress on short-horizon, well-scoped tasks, yet their ability to sustain coherent decisions in d

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Scalable and Trustworthy Earth Observation Foundation Models

DGX agent

arXiv:2607.07758v1 Announce Type: new Abstract: Foundation models (FMs) have transformed machine learning from isolated task-specific model development toward general-purpose models pretrained on broa

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Secure Decentralized Federated Learning via Gossip and Virtual Voting

DGX agent

arXiv:2607.08651v1 Announce Type: new Abstract: Decentralized federated learning (DFL) removes the central server by letting nodes exchange model updates through peer-to-peer gossip, but existing goss

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Shift & Drift: A Zero-Shot Benchmark for Generalizable and Robust Autonomous Driving Motion Planning

DGX agent

arXiv:2607.07844v1 Announce Type: cross Abstract: While closed-loop motion planners trained on large-scale, object-level datasets, e.g., nuPlan, demonstrate strong in-distribution (ID) performance, th

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets

DGX agent

arXiv:2607.08681v1 Announce Type: new Abstract: As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trustwo

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

SQuaD-SQL: Efficient Text-to-SQL with Small Language Models via LLM-Guided Knowledge Distillation

DGX agent

arXiv:2607.08161v1 Announce Type: new Abstract: Text-to-SQL is a fundamental task in natural language processing that enables users to interact with structured databases using natural language. While

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Structured Pruning of Large Language Models via Power Transformation and Sign-Preserving Score Aggregation with Adaptive Feature Retention

DGX agent

arXiv:2607.08027v1 Announce Type: cross Abstract: This paper proposes an improved structured pruning method for large language models (LLMs) that addresses key challenges in adapting Adaptive Feature

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Super Weights in LLMs and the Failure of Selective Training

DGX agent

arXiv:2607.08733v1 Announce Type: new Abstract: Recent work identified Super Weights, individual parameters whose removal degrades model performance by orders of magnitude. We show that this degradati

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

SwinIFS: Landmark Guided Swin Transformer For Identity Preserving Face Super Resolution

DGX agent

arXiv:2601.01406v2 Announce Type: replace-cross Abstract: Face super-resolution aims to recover high-quality facial images from severely degraded low-resolution inputs, but remains challenging due to

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

TFP: Temporally Conditioned Memory-Fusion Policies for Visuomotor Learning

DGX agent

arXiv:2607.08283v1 Announce Type: new Abstract: Vision--Language--Action (VLA) policies such as pi_{0.5} and OpenVLA perform well on many manipulation tasks, but they are often reactive: the next acti

model-releasesarxiv-cs-ro
10 Jul 2026
Model Releases

The Phasor Transformer: Resolving Attention Bottlenecks on the Unit Circle

DGX agent

arXiv:2603.17433v2 Announce Type: replace-cross Abstract: Transformer models have redefined sequence learning, yet dot-product self-attention introduces a quadratic token-mixing bottleneck for long-co

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

The Regularization Parameter: Sparse Precision Matrix Estimation

DGX agent

arXiv:2607.07735v1 Announce Type: cross Abstract: Sparse precision matrix estimation provides an interpretable and computationally efficient framework for modeling conditional dependencies in high-dim

model-releasesarxiv-cs-lg
10 Jul 2026
← Previous
1…7677787980…361
Next →