AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Empathy Applicability Modeling for General Health Queries

DGX agent

arXiv:2601.09696v2 Announce Type: replace Abstract: LLMs are increasingly being integrated into clinical workflows, yet they often lack clinical empathy, an essential aspect of effective doctor-patien

model-releasesarxiv-cs-cl
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Enhancing Blind Source Separation with Dissociative Principal Component Analysis

DGX agent

arXiv:2411.12321v2 Announce Type: replace Abstract: Principal component analysis (PCA) and its sparse variants (sPCA) are widely used as a precursor to independent component analysis (ICA) for blind s

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Enhancing LLM Metacognition via Cognitive Pairwise Training

DGX agent

arXiv:2606.00869v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become central to LLM reasoning, but its outcome-level rewards can make models more willing to

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Error Bounds for a Diffusion Model-Based Drift Estimator

DGX agent

arXiv:2606.02115v1 Announce Type: cross Abstract: Parameter estimation in stochastic differential equations is a classical statistical problem of much importance in many scientific fields. Recent work

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

ES-Merging: Biological MLLM Merging via Embedding Space Signals

DGX agent

arXiv:2603.14405v2 Announce Type: replace-cross Abstract: Biological multimodal large language models (MLLMs) have emerged as powerful foundation models for scientific discovery. However, existing mod

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding

DGX agent

arXiv:2603.03312v3 Announce Type: replace-cross Abstract: Decoding natural language from non-invasive EEG signals is a promising yet challenging task. However, current state-of-the-art models remain c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Escaping the Mode Lottery: Multi-Response Training Improves Language Model Generalization

DGX agent

arXiv:2606.00544v1 Announce Type: cross Abstract: Modern language-model fine-tuning typically pairs each prompt with a single response, even though many prompts admit multiple valid completions. This

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

EuraGovExam: A Multilingual Multimodal Benchmark from Real-World Civil Service Exams

DGX agent

arXiv:2603.27223v2 Announce Type: replace-cross Abstract: We present EuraGovExam, a multilingual and multimodal benchmark sourced from real-world civil service examinations across five representative

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating Interactive Reasoning in Large Language Models: A Hierarchical Benchmark with Executable Games

DGX agent

arXiv:2606.00103v1 Announce Type: new Abstract: We introduce a multi-turn interactive framework for reasoning evaluation that treats reasoning as active evidence acquisition and belief updating. Where

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating Real-World Generalizability of Algorithm Selection Models

DGX agent

arXiv:2606.02016v1 Announce Type: new Abstract: Algorithm Selection (AS) aims to automatically identify the most suitable optimization algorithm for a given problem instance by leveraging measurable p

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers

DGX agent

arXiv:2602.22221v2 Announce Type: replace-cross Abstract: Search engines and AI-powered systems increasingly mediate access to factual information, yet their reliability remains difficult to evaluate

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Evaluating the Reversal Curse in Model Editing

DGX agent

arXiv:2310.10322v3 Announce Type: replace Abstract: Large language models (LLMs) are prone to hallucinate unintended text due to false or outdated knowledge. Since retraining LLMs is resource intensiv

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Explainable Forensics of Manipulated Segments in Untrimmed Long Videos

DGX agent

arXiv:2606.02402v1 Announce Type: new Abstract: The rapid advancement of AI-driven video generation has transformed content creation, while simultaneously increasing the risk of misinformation through

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression

DGX agent

arXiv:2606.01608v1 Announce Type: new Abstract: Most existing extreme compression methods fail to achieve an optimal rate-distortion-perception trade-off, as they typically prioritize perceptual fidel

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays

DGX agent

arXiv:2509.15234v2 Announce Type: replace Abstract: Multimodal learning from paired medical images and clinical text is a central challenge in medical data-driven informatics, where effective cross-mo

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ExpWeaver: LLM Agents Learn from Experience via Latent RAG

DGX agent

arXiv:2606.01041v1 Announce Type: new Abstract: Experience learning has achieved promising results in enhancing LLM agent planning and reasoning by integrating past interactions as reusable knowledge.

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

FACT: A Simple and Efficient Framework for Active Finetuning

DGX agent

arXiv:2606.02079v1 Announce Type: new Abstract: The main goal of active finetuning is to improve a pretrained model's performance on a specific task or domain by finetuning it with carefully selected

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

FALAT: Tracing Failures in LLM Agent Trajectories via Dependency-Guided Search

DGX agent

arXiv:2606.00765v1 Announce Type: new Abstract: LLM-based agents increasingly solve complex tasks through long trajectories involving reasoning steps, tool calls, and inter-agent communication. Howeve

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Fast-SAM3D: 3Dfy Anything in Images but Faster

DGX agent

arXiv:2602.05293v2 Announce Type: replace Abstract: SAM3D enables scalable, open-world 3D reconstruction from complex scenes, yet its deployment is hindered by prohibitive inference latency. In this w

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Faster Synchronous On-Policy RL via Straggler-Aware Group Sizing

DGX agent

arXiv:2606.02218v1 Announce Type: cross Abstract: Synchronous reinforcement learning methods such as Group Relative Policy Optimization (GRPO) provide stable and reproducible on-policy training, but t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Feature to Dynamics: Feature-space to Autoregression strategy for Zero-shot Time Series Forecasting

DGX agent

arXiv:2606.01289v1 Announce Type: new Abstract: Zero-shot time series forecasting aims to predict future values for previously unseen series, requiring models to generalize temporal dynamics beyond th

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning

DGX agent

arXiv:2604.03893v2 Announce Type: replace Abstract: Current multimodal benchmarks for scientific reasoning primarily evaluate local information extraction -- models recognize symbols and values and th

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes

DGX agent

arXiv:2606.02523v1 Announce Type: new Abstract: Suicide memes are memes used to express suicide-related thoughts or comment on suicide-related issues. Suicide memes are increasingly common on social m

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Finding the Minimal Parameter Budget for Implicit Reasoning: A Data Complexity Driven Scaling Law for Language Models

DGX agent

arXiv:2504.03635v4 Announce Type: replace Abstract: Reasoning is a core capability of language models (LMs), yet it remains unclear how much model capacity is necessary to support reasoning during pre

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Fine-Tuning Diffusion Models for Molecular Generation via Reinforcement Learning and Fast Sampling

DGX agent

arXiv:2606.01220v1 Announce Type: cross Abstract: Generating molecules that simultaneously satisfy drug-like properties and conform to the 3D structure of a target protein is a core challenge in struc

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Finer Parameter Steps for Low-Rank PEFT: A Controlled Study with CP Tensor Adapters

DGX agent

arXiv:2606.00428v1 Announce Type: cross Abstract: Low-rank adapters are usually compared by sweeping a small set of ranks, but the rank also fixes the resolution of the parameter budget. For a 2048{im

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

FineVerify: Scaling Test-Time Compute with Fine-Grained Self-Verification for Agentic Search

DGX agent

arXiv:2606.00660v1 Announce Type: new Abstract: Agentic search requires language model agents to explore many sources and answer complex information-seeking questions. Scaling test-time compute is a p

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

FLAME: Physics-Guided Neural Operators for Onboard Satellite Methane Detection in Hyperspectral Imagery

DGX agent

arXiv:2606.01577v1 Announce Type: new Abstract: Methane is a major driver of near-term climate change, and rapidly identifying its emission sources is a critical climate intervention. Spaceborne hyper

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Flow Matching for Convective-Scale Precipitation Downscaling

DGX agent

arXiv:2606.00281v1 Announce Type: cross Abstract: Generative machine learning is an increasingly important complement to dynamical downscaling for producing high-resolution precipitation projections,

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Flowers: A Warp Drive for Neural PDE Solvers

DGX agent

arXiv:2603.04430v2 Announce Type: replace Abstract: We introduce Flowers, a neural architecture for learning PDE solution operators built entirely from multihead warps. Aside from pointwise channel mi

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

FlowIt: Global Matching via Hierarchical Transformers and Optimal Transport for Optical Flow

DGX agent

arXiv:2603.28759v2 Announce Type: replace Abstract: We present FlowIt, a novel architecture for optical flow estimation that combines global matching with confidence and occlusion-guided refinement. A

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection

DGX agent

arXiv:2606.00782v1 Announce Type: new Abstract: Open-vocabulary object detection (OVD) has achieved remarkable progress through large-scale vision-language pre-training. Existing methods, however, typ

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment

DGX agent

arXiv:2606.00644v1 Announce Type: new Abstract: AI research often requires decisions before future evidence exists: which bottleneck to attack, which direction to pursue, or where a project should be

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

From Evaluation to Design: Using Potential Energy Surface Smoothness Metrics to Guide Machine Learning Interatomic Potential Architectures

DGX agent

arXiv:2602.04861v2 Announce Type: replace-cross Abstract: Machine Learning Interatomic Potentials (MLIPs) sometimes fail to reproduce the physical smoothness of the quantum potential energy surface (P

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

From Moments to Models: Graphon-Mixture Learning for Mixup and Contrastive Learning

DGX agent

arXiv:2510.03690v4 Announce Type: replace Abstract: Real-world graph datasets often arise from mixtures of populations, where graphs are generated by multiple distinct underlying distributions. In thi

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

From Outliers to Errors: Auditing Pali-to-English LLM Translations with Multi-Reference Adjudication

DGX agent

arXiv:2606.01136v1 Announce Type: new Abstract: Single-score translation metrics can conflate legitimate variation with error, a problem especially acute for classical languages where multiple defensi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

From Scaling to Structured Expressivity: Rethinking Transformers for CTR Prediction

DGX agent

arXiv:2511.12081v2 Announce Type: replace-cross Abstract: Despite massive investments in scale, deep models for click-through rate (CTR) prediction often exhibit rapidly diminishing returns -- a stark

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

From Segments to Scenes: Temporal Understanding in Autonomous Driving via Vision-Language Model

DGX agent

arXiv:2512.05277v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as the perception and reasoning backbone of autonomous agents acting in the wild, with

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

From Unfamiliar to Familiar: Detecting Pre-training Data via Gradient Deviations in Large Language Models

DGX agent

arXiv:2603.04828v2 Announce Type: replace Abstract: Pre-training data detection for LLMs is essential for addressing copyright concerns and mitigating benchmark contamination. Existing methods mainly

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

FVSpec: Real-World Property-Based Tests as Lean Challenges

DGX agent

arXiv:2606.01008v1 Announce Type: cross Abstract: We present a benchmark for evaluating AI models and agents on real-world formal software verification tasks. We first scrape 11,039 property-based tes

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

G2LoRA: Gradient Orthogonal Low-Rank Adaptation Framework for Graph Continual Learning on Text-Attributed Graphs

DGX agent

arXiv:2606.01873v1 Announce Type: new Abstract: LLM-as-Aligner has emerged as a prevalent pre-training paradigm for Text-Attributed Graphs(TAGS), aligning graph and text modalities into a shared embed

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

GABI: Geometry-Aware Boundary Integration for Spacecraft Segmentation

DGX agent

arXiv:2606.00886v1 Announce Type: new Abstract: Accurate segmentation is crucial for autonomous spacecraft, as it directly affects downstream tasks related to 3D situational awareness. The harsh illum

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

GenPT: Beyond Self-Report for Reliable LLM Psychometrics via Generative Projective Testing

DGX agent

arXiv:2606.00860v1 Announce Type: cross Abstract: Self-report questionnaires remain the prevailing tool for probing the psychological states of persona-conditioned agents (PC-Agents). However, classic

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Geometry-Aware Implicit Memory for Video World Models

DGX agent

arXiv:2606.02436v1 Announce Type: new Abstract: Video world models aim to simulate controllable visual environments, but long-horizon rollouts depend on what the model remembers after observations lea

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

GIRL-DETR: Gradient-Isolated Reinforcement Learning for Video Moment Retrieval

DGX agent

arXiv:2606.00775v1 Announce Type: cross Abstract: Video Moment Retrieval (VMR) task requires accurately localizing temporal boundaries aligned with natural language queries, but many models suffer fro

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

GLENS: Global Search via Learning from Solver Iterates with Diffusion Models

DGX agent

arXiv:2606.00366v1 Announce Type: new Abstract: We consider the problem of generating a large collection of initial guesses for local minima of multimodal non-convex continuous optimization problems.

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures

DGX agent

arXiv:2510.24081v2 Announce Type: replace Abstract: To date, there exist almost no culturally-specific evaluation benchmarks for large language models (LLMs) that cover a large number of languages and

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training

DGX agent

arXiv:2606.00539v1 Announce Type: new Abstract: Training stability is a key bottleneck in low-precision language model training: efficient low-cost paths can still produce short-lived numerical risks

model-releasesarxiv-cs-lg
2 Jun 2026
← Previous
1…166167168169170…361
Next →