AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,514 results
Model Releases

AirGroundBench: Probing Spatial Intelligence in Multimodal Large Models under Heterogeneous Multi-View Embodied Collaboration

DGX agent

arXiv:2606.28049v1 Announce Type: new Abstract: In recent years, multimodal large language models (MLLMs) have shown strong potential for embodied intelligence, yet their ability to maintain geometric

model-releasesarxiv-cs-cv
29 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CalBrief: A Pilot Diagnostic Benchmark for Evidence-Calibrated Scientific Briefing with Large Language Models

DGX agent

arXiv:2606.27383v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as research assistants, yet it remains unclear whether they can calibrate research takeaways to the

model-releasesarxiv-cs-ai
29 Jun 2026
Local Ai

Continual Learning for Sequential Personalization of Small Language Models: A Stability Monitoring Analysis

DGX agent

arXiv:2606.27634v1 Announce Type: new Abstract: Small Language Models (SLMs) are increasingly being considered for deployment on edge devices such as laptops, enabling private, low-latency, and locall

local-aiarxiv-cs-lg
29 Jun 2026
Model Releases

DeepSeek details DSpark, a speculative decoding framework for its V4 models, saying it speeds up AI inference by up to 85% and was tested on Gemma and Qwen (Ben Jiang/South China Morning Post)

DGX agent

Ben Jiang / South China Morning Post: DeepSeek details DSpark, a speculative decoding framework for its V4 models, saying it speeds up AI inference by up to 85% and was tested on Gemma and Qwen — Chin

model-releasestechmeme
29 Jun 2026
Applications

Foundation vs. Specialized Models: Evaluating Catastrophic Forgetting in Continual Time Series Forecasting

DGX agent

arXiv:2510.00809v3 Announce Type: replace Abstract: While Time Series Foundation Models (TSFMs) excel in zero-shot tasks, their behavior under continual fine tuning is poorly understood. We present th

applicationsarxiv-cs-lg
29 Jun 2026
Model Releases

From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models

DGX agent

arXiv:2606.27679v1 Announce Type: cross Abstract: Probe-based uncertainty estimation (UE) has emerged as a prominent approach to detect hallucinations in Large Language Models (LLMs) by learning uncer

model-releasesarxiv-cs-ai
29 Jun 2026
Research

Hippocampus-DETR: An Explicit Memory Object Detection Framework Based on Hippocampus Modeling

DGX agent

arXiv:2606.27831v1 Announce Type: cross Abstract: This paper addresses the lack of explicit memory mechanisms in current object detection models and proposes Hippocampus-DETR, a novel detection framew

researcharxiv-cs-ai
29 Jun 2026
Research

Large Language Model Teaches Visual Students: Cross-Modality Transfer of Fine-Grained Conceptual Knowledge

DGX agent

arXiv:2606.27527v1 Announce Type: cross Abstract: Large Language Models (LLMs) possess broad conceptual knowledge acquired through large-scale text pretraining, yet their potential to supervise models

researcharxiv-cs-ai
29 Jun 2026
Model Releases

MemoBench: Benchmarking World Modeling in Dynamically Changing Environments

DGX agent

arXiv:2606.27537v1 Announce Type: new Abstract: Video generation models aspire to simulate dynamic environments, and several benchmarks now evaluate memory consistency across frames. However, most ass

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving

DGX agent

arXiv:2606.27660v1 Announce Type: new Abstract: Vision-Language Models (VLMs) improve generalization and interpretability in autonomous driving but suffer from efficiency issues due to long visual tok

model-releasesarxiv-cs-cv
29 Jun 2026
Tutorials

RAE-NWM: Navigation World Model in Dense Visual Representation Space

DGX agent

arXiv:2603.09241v2 Announce Type: replace Abstract: Visual navigation requires agents to reach goals in complex environments through perception and planning. World models address this task by simulati

tutorialsarxiv-cs-cv
29 Jun 2026
Safety

RECAST: Model Reconstruction via Counterfactual-Aware Wasserstein Geometry under Limited Data

DGX agent

arXiv:2606.27948v1 Announce Type: new Abstract: Counterfactual explanations (CFs) help understand machine learning models by identifying minimal input changes that would lead to alternative model outc

safetyarxiv-cs-lg
29 Jun 2026
Model Releases

SemCityLoc: Aerial 6DoF Localization Using Semantic 3D City Models

DGX agent

arXiv:2606.27444v1 Announce Type: new Abstract: Aerial 6DoF localization typically relies on precise GNSS signals or radiometrically rich 3D reconstructions, limiting scalability and on-board deployme

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

SurgXBench: Explainable Vision-Language Model Benchmark for Surgery

DGX agent

arXiv:2505.10764v4 Announce Type: replace Abstract: Innovations in digital intelligence are transforming robotic surgery with more informed decision-making. Real-time awareness of surgical instrument

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

This is smart from Cline. They just launched ClinePass, which makes it easy to access the latest open-weight models like GLM 5.2, Kimi k2.7-…

DGX agent

This is smart from Cline. They just launched ClinePass, which makes it easy to access the latest open-weight models like GLM 5.2, Kimi k2.7-code, Mimo 2.5, Deepseek v4 pro, Minimax M3, and more. Alway

model-releasesdair-ai--x
29 Jun 2026
Industry

We have seen multi model harnesses for cheaper & faster tasks What about for the hardest challenges? What about open source? Proud to share …

DGX agent

We have seen multi model harnesses for cheaper & faster tasks What about for the hardest challenges? What about open source? Proud to share the latest update our Zenith harness, taking models you can

industryemad-mostaque--x
29 Jun 2026
Model Releases

Sakana Fugu Technical Report Instead of training one larger model, Sakana AI trains an orchestrator that reads each query and dynamically ro…

DGX agent

Sakana Fugu Technical Report Instead of training one larger model, Sakana AI trains an orchestrator that reads each query and dynamically routes or composes GPT-5.5, Gemini-3.1-Pro, Claude Opus 4.8 an

model-releasesdavid-ha--x
27 Jun 2026
Model Releases

Can Large Language Models Reliably Code Qualitative Humanitarian Data? A Benchmark Study Against Human Expert Adjudication

DGX agent

arXiv:2606.26541v1 Announce Type: new Abstract: Data from affected populations are crucial for informing humanitarian response, but their value depends on timely and consistent interpretation of nuanc

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

Don't Settle at the Mode! Mitigating Diversity Collapse in Pretrained Flow Models via Feature Self-Guidance

DGX agent

arXiv:2606.27371v1 Announce Type: new Abstract: State-of-the-art flow models generate stunning images from text or image prompts. However, they suffer from diversity collapse when generating multiple

safetyarxiv-cs-cv
26 Jun 2026
Model Releases

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models

DGX agent

arXiv:2606.26196v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have recently made remarkable progress in unifying vision-language understanding and reasoning, especially fo

model-releasesarxiv-cs-ai
26 Jun 2026
Local Ai

Heterogeneous Neural Predictivity from Language Models During Naturalistic Comprehension

DGX agent

arXiv:2606.26880v1 Announce Type: new Abstract: Language-model representations provide structured, high-dimensional annotations of naturalistic language stimuli and can serve as informative neural pre

local-aiarxiv-cs-cl
26 Jun 2026
Model Releases

NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Language Models

DGX agent

arXiv:2606.27047v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance across a wide range of tasks, but ensuring their reliability in highly technical dom

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Refusal Lives Downstream of Persona in Chat Models

DGX agent

arXiv:2606.26161v1 Announce Type: new Abstract: Linear directions in activation space have been identified for both refusal and persona traits in instruction-tuned chat models, but the two have been s

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning

DGX agent

arXiv:2606.26290v1 Announce Type: cross Abstract: While parameter-efficient fine-tuning (PEFT) typically targets attention projectors, its efficacy for tasks requiring sequential state accumulation re

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Unconventional AI debuts oscillator-based Un-0 model series

DGX agent

Unconventional AI Inc. has developed an artificial intelligence architecture that could improve the power efficiency of image generation models. The technology is the basis of a new neural network ser

model-releasessiliconangle
26 Jun 2026
Research

Where Larger Models Excel: The Primacy of Constraint-Guided Reasoning

DGX agent

arXiv:2606.26108v1 Announce Type: new Abstract: Larger language models consistently outperform smaller ones on reasoning benchmarks, yet the reasoning differences underlying this gap remain underexplo

researcharxiv-cs-cl
26 Jun 2026
Safety

Adaptive Oscillatory Inductive Bias for Modeling Sharp Prosodic Dynamics in Diffusion-Based TTS

DGX agent

arXiv:2606.25424v1 Announce Type: cross Abstract: Diffusion-based text-to-speech (TTS) models have achieved significant improvements in speech quality. However, modeling sharp prosodic transitions and

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Agent-as-a-Router: Agentic Model Routing for Coding Tasks

DGX agent

arXiv:2606.22902v2 Announce Type: replace Abstract: Real-world users typically have access to multiple Large Language Models (LLMs) from different providers, and these LLMs often excel at distinct dom

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

Beyond Next-Observation Prediction: Agent-Authored World Modeling for Sequential Decision Making

DGX agent

arXiv:2606.25421v1 Announce Type: new Abstract: Recent studies on world modeling for Large Language Model (LLM) agents typically formulate the learning objective as next-observation prediction. Howeve

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

How Small Can 6G Reason? Scaling Tiny-to-Small Language Models for AI-Native Networks

DGX agent

arXiv:2603.02156v2 Announce Type: replace-cross Abstract: Emerging 6G visions, reflected in ongoing standardization efforts within 3GPP, IETF, ETSI, ITU-T, and the O-RAN Alliance, increasingly charact

model-releasesarxiv-cs-ai
25 Jun 2026
Research

MIMFlow: Integrating Masked Image Modeling with Normalizing Flows for End-to-End Image Generation

DGX agent

arXiv:2606.26016v1 Announce Type: new Abstract: Normalizing Flows (NFs) are powerful generative models capable of exact density estimation and sampling. However, their strict invertibility often force

researcharxiv-cs-cv
25 Jun 2026
Model Releases

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models

DGX agent

arXiv:2606.25442v1 Announce Type: new Abstract: Safety alignment of large language models (LLMs) typically depends on high-quality supervision data, such as safe demonstrations or preference pairs. Ho

model-releasesarxiv-cs-cl
25 Jun 2026
Research

RotRNN: Modelling Long Sequences with Rotations

DGX agent

arXiv:2407.07239v3 Announce Type: replace Abstract: Linear recurrent neural networks, such as State Space Models (SSMs) and Linear Recurrent Units (LRUs), have recently shown state-of-the-art performa

researcharxiv-cs-lg
25 Jun 2026
Model Releases

The underrated part of this announcement is that Fireworks has been quietly great behind the scenes helping us eval and serve these models B…

DGX agent

The underrated part of this announcement is that Fireworks has been quietly great behind the scenes helping us eval and serve these models Big kudos to the Fireworks team! Kimi K2.7 Code and GLM 5.2 a

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

Toward Low-Latency Vision-Language Models with Doubly-Correct Predictions in Egocentric Visual Understanding

DGX agent

arXiv:2606.25160v1 Announce Type: cross Abstract: The rapid rise of Vision-Language Models (VLMs) in egocentric visual understanding has made low-latency inference in human-robot collaborative (HRC) t

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

When Do Conservation Laws Survive Learned Representations? Certified Horizons for Latent World Models

DGX agent

arXiv:2606.24945v1 Announce Type: new Abstract: We ask a representation-learning question about physical world models: when does a conservation law remain certifiable after a model learns a latent rep

safetyarxiv-cs-lg
25 Jun 2026
Safety

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety

DGX agent

arXiv:2606.25034v1 Announce Type: new Abstract: General-purpose models often struggle to reliably identify and understand real-world multimodal risks, largely due to the inherent multimodal adversaria

safetyarxiv-cs-cv
25 Jun 2026
Model Releases

A Physics-Informed Fourier-Wavelet Transformer for Multiscale Computational Fluid Dynamics Surrogate Modeling

DGX agent

arXiv:2606.24696v1 Announce Type: cross Abstract: Physics-informed surrogate models can accelerate computational fluid dynamics simulations. However, many existing methods reproduce global flow patter

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

A specialized reasoning large language model for accelerating rare disease diagnosis: a randomized AI physician assistance trial

DGX agent

arXiv:2606.24510v1 Announce Type: new Abstract: Rare diseases affect millions of individuals worldwide, yet timely diagnosis remains a major public health challenge due to scarcity of specialized clin

model-releasesarxiv-cs-ai
24 Jun 2026
Research

DREAM: Dense Retrieval Embeddings via Autoregressive Modeling

DGX agent

arXiv:2606.24667v1 Announce Type: new Abstract: Dense retrieval embedding models are a fundamental component of modern retrieval-based AI systems. Most dense retrievers are trained with contrastive ob

researcharxiv-cs-cl
24 Jun 2026
Research

Flood Mapping from RGB imagery using a Vision Foundation Model

DGX agent

arXiv:2606.24120v1 Announce Type: new Abstract: Timely, high-resolution maps of flood extent around settlements are essential for emergency response and damage assessment. We consider airborne RGB ima

researcharxiv-cs-cv
24 Jun 2026
Model Releases

GeoT2V-Bench: Benchmarking 3D Consistency in Text-to-Video Models via 3D Reconstruction

DGX agent

arXiv:2606.24829v1 Announce Type: new Abstract: Camera-prompted text-to-video (T2V) models are increasingly used to synthesize virtual camera captures, such as orbiting objects or moving through stati

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

GLM-5.2 is now available in Cursor. The model has performed strongly on OpenRouter's Cursor usage rankings over the past week. Would love to…

DGX agent

GLM-5.2 is now available in Cursor. The model has performed strongly on OpenRouter's Cursor usage rankings over the past week. Would love to hear comparisons of the experience using BYOK (GLM Coding P

model-releaseszhipu-ai--x
24 Jun 2026
Model Releases

Performance and Interpretability of Convolutional, Transformer, and Hybrid Deep Learning Models in Colorectal Histology Classification

DGX agent

arXiv:2606.23744v1 Announce Type: cross Abstract: Deep learning has become an important tool in computational pathology, enabling automated analysis of histopathological images. While convolutional ne

model-releasesarxiv-cs-cv
24 Jun 2026
Research

PORTER: Language-Grounded Event Representations for Portable Structured EHR Foundation Models

DGX agent

arXiv:2606.24102v1 Announce Type: new Abstract: Most electronic health record (EHR) foundation models encode clinical events as discrete event tokens from a fixed vocabulary and therefore cannot direc

researcharxiv-cs-cl
24 Jun 2026
Model Releases

RetiSEM: Generalising Causal Models for Fragmented Biomedical Data

DGX agent

arXiv:2606.24488v1 Announce Type: cross Abstract: Learning causal models from fragmented biomedical data is challenging because clinical, molecular, and imaging variables are often incomplete or not j

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling

DGX agent

arXiv:2606.24187v1 Announce Type: new Abstract: Long video understanding remains a daunting challenge for Multimodal Large Language Models (MLLMs) due to the excessive computation and memory footprint

model-releasesarxiv-cs-cv
24 Jun 2026
Applications

A polarity-aware multi-relational model for the signed interaction prediction in biological networks

DGX agent

arXiv:2407.07357v4 Announce Type: replace Abstract: Predicting signed interactions in biological networks is crucial for understanding drug mechanisms and facilitating drug repurposing. While deep gra

applicationsarxiv-cs-lg
23 Jun 2026
← Previous
1…103104105106107…1261
Next →