AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,503 results
Model Releases

DASH: Divergence-Adaptive Supervision Horizons for On-Policy Self-Distillation of Reasoning Models

DGX agent

arXiv:2608.06243v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models using automatically verifiable outcom

model-releasesarxiv-cs-ai
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DynaPix: Can Vision-Language Models Identify the Exact Future?

DGX agent

arXiv:2608.05505v1 Announce Type: new Abstract: Acting in a physical scene requires knowing its real later state, not a plausible one. Current evaluations often accept words or a realistic-looking ima

model-releasesarxiv-cs-cv
7 Aug 2026
Research

Evaluating Machine Learning Models for Post-Wildfire Debris-Flow Prediction

DGX agent

arXiv:2608.05265v1 Announce Type: cross Abstract: Prediction of post-wildfire debris flows is critical for mitigating hazards to communities, infrastructure, and resources during intense rainfall in r

researcharxiv-cs-cv
7 Aug 2026
Safety

GeniWorld: A Generalizable Interactive World Model for Robotic Manipulation via Visual Actions

DGX agent

arXiv:2608.06332v1 Announce Type: new Abstract: Generalist robot policies exhibit strong capabilities, but their robustness in complex and unseen environments remains limited. Scaling robot learning a

safetyarxiv-cs-ro
7 Aug 2026
Model Releases

Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations

DGX agent

arXiv:2604.17359v2 Announce Type: replace-cross Abstract: Language models asked to simulate psychiatric patients produce cases that survive inspection one at a time and populations that match no real

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

An accurate characterization of the arc of AI is that it is shaped by two trends: 1. Moving more and more logic to a neural model for tasks …

DGX agent

An accurate characterization of the arc of AI is that it is shaped by two trends: 1. Moving more and more logic to a neural model for tasks where training data can be densely sampled (e.g. the shift f

agentsfrancois-chollet--x
6 Aug 2026
Safety

CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models

DGX agent

arXiv:2608.04302v1 Announce Type: new Abstract: Benchmarking video-language models has largely focused on short clips and single-sentence metrics, leaving open whether current systems can generate acc

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

Faster-WAM: Efficient Inference-Time Future Conditioning for Robust World Action Models

DGX agent

arXiv:2608.04404v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot manipulation by learning how the environment evolves beyond the current observation. However, existing approach

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

From Financial Sentiment Classification to Return Predictability: A QLoRA Benchmark of Large Language Models

DGX agent

arXiv:2608.04200v1 Announce Type: cross Abstract: Financial sentiment classifiers are commonly evaluated against human labels, but strong linguistic performance does not necessarily imply economically

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework for Large Audio-Language Models

DGX agent

arXiv:2608.04351v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) excel at general speech understanding; however, adapting them to fine-grained tasks like Speech Emotion Recognitio

model-releasesarxiv-cs-ai
6 Aug 2026
Local Ai

i just spent weeks rewriting my webUI from scratch, getting rid of all AI slop within the codebase and switching it over to a proper lightweight framework (alpine.js). i am now comfortable suggesting it as an alternative to openwebUI, librechat and the like! it is made for local models

DGX agent

[Fully open source under GPL3, made from the ground up for use with local models, no subscriptions, no corporate backing] When i first started this, it was meant to be a fully lightweight, extremely m

local-air-localllama
6 Aug 2026
Model Releases

LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling

DGX agent

arXiv:2608.04147v1 Announce Type: cross Abstract: Label noise is common in medical imaging datasets due to factors such as inter-rater variability, annotation errors, and ambiguous cases. This can sev

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Mamba with Hierarchical Memory: Solving Representation Bottleneck in Long Sequence Modeling

DGX agent

arXiv:2608.02347v2 Announce Type: replace Abstract: Recurrent linear attention models (RLAs) such as Mamba offer efficient linear-time sequence modeling as an alternative to Transformers, yet their fi

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Same Formulas, Different Semantics: Do Language Models Follow Modal Logic Specifications?

DGX agent

arXiv:2608.05097v1 Announce Type: new Abstract: Reasoning about necessity and possibility depends on assumptions about accessibility between worlds and about which objects exist at each one. The same

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Teaching Foundation Models to Read mmWave: Pose-Guided Kinematic Representation for Human Behavior Understanding

DGX agent

arXiv:2608.04127v1 Announce Type: new Abstract: Large language model agents need to perceive human behavior in physical environments. Millimeter-wave (mmWave) radar provides a privacy-friendly and con

model-releasesarxiv-cs-cv
6 Aug 2026
Safety

A Blind Spot in Alignment: Quantifying Biosecurity Risks in Large Language Models

DGX agent

arXiv:2608.02684v1 Announce Type: cross Abstract: Large Language Models (LLMs) are accelerating biological research, yet this same capability poses a critical biosecurity threat: models that assist in

safetyarxiv-cs-ai
5 Aug 2026
Safety

A game theory for foundation models shows new paths to rational cooperation through similarity inference

DGX agent

arXiv:2608.03958v1 Announce Type: new Abstract: As autonomous agents powered by foundation models are increasingly integrated into social and economic systems, understanding the principles governing t

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Can Text-to-Image Models Draw from the Right Frame of Reference?

DGX agent

arXiv:2608.03357v1 Announce Type: new Abstract: Spatial instruction following has become a crucial requirement for text-to-image (T2I) generation. A common challenge arises when directional expression

model-releasesarxiv-cs-cv
5 Aug 2026
Research

Caved or Convinced: Temporal Sampling Gates Claim Deference in Video Large Language Models

DGX agent

arXiv:2608.03160v1 Announce Type: cross Abstract: When asked which of two events came first, video large language models can fail in two opposite ways: cave to a false claim, or reject a true one. Pri

researcharxiv-cs-cv
5 Aug 2026
Model Releases

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili

DGX agent

arXiv:2608.03532v1 Announce Type: new Abstract: Large language models are increasingly deployed in multilingual contexts, yet safety alignment and bias evaluation remain overwhelmingly English-centric

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

dots.tts.edit: Precisely Controlled Speech Editing with a Continuous Autoregressive Model

DGX agent

arXiv:2608.02673v1 Announce Type: cross Abstract: Speech editing for content creation requires precise control over both what an edit should do and where it should apply. Free-form natural language pr

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

HomeSafeBench: A Benchmark for Embodied Vision-Language Models in Free-Exploration Home Safety Inspection

DGX agent

arXiv:2509.23690v2 Announce Type: replace-cross Abstract: Safety hazards in the home are a leading cause of preventable domestic injuries, motivating an automated inspector that actively explores a ho

model-releasesarxiv-cs-cl
5 Aug 2026
Hardware

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

DGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

hardwarearxiv-cs-ai
5 Aug 2026
Research

MDLMPE: Distribution Aware Positional Encoding for Masked Diffusion Language Models

DGX agent

arXiv:2608.03769v1 Announce Type: cross Abstract: Masked diffusion language models (MDLMs) enable parallel generation and bidirectional context modeling, but their positional context differs fundament

researcharxiv-cs-ai
5 Aug 2026
Model Releases

MIDI-LLM: Improving Text-to-MIDI Music Generation via Adapting Large Language Models

DGX agent

arXiv:2511.03942v2 Announce Type: replace-cross Abstract: We present MIDI-LLM, a recipe that improves multitrack text-to-MIDI generation via adapting Large Language Models (LLMs). MIDI-LLM expands an

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Qwen-3D: A Generalist 3D Vision-Language Model for Spatial Understanding

DGX agent

arXiv:2608.02980v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have achieved remarkable success on images and short videos, yet scaling them to long videos remains challenging due to f

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

Self-Guided Adaptive Safety Alignment: Synthesizing and Internalizing Guidelines in Reasoning Models

DGX agent

arXiv:2511.21214v4 Announce Type: replace-cross Abstract: Explicit safety policies can improve reasoning-model safety, but their effective coverage may lag behind evolving jailbreak strategies. We stu

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

SeqLLM: Augmenting LLMs with Behavioral-Sequence Modeling for High-Stakes Decisions at WeChat Pay

DGX agent

arXiv:2608.03063v1 Announce Type: new Abstract: Merchant risk control at large payment platforms screens tens of millions of merchants daily, where false positives harm legitimate merchants and false

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SlimVLM: Sensitivity-aware Dynamic Structured Pruning with Adaptive Visual Token Selection for Efficient Vision-Language Models

DGX agent

arXiv:2608.03580v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have demonstrated remarkable performance in processing and understanding both text and images, their large parameter

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling

DGX agent

arXiv:2608.02951v1 Announce Type: cross Abstract: Preference-based reinforcement learning (PbRL) for general stochastic MDPs often requires training a reward model. Existing reward-model-free methods

safetyarxiv-cs-ai
5 Aug 2026
Safety

A Heuristic Perspective on Debiasing Language Models

DGX agent

arXiv:2608.00622v1 Announce Type: new Abstract: Language models (LMs) often acquire various biases during pre-training and may express them in interactions, potentially causing social harm. Existing m

safetyarxiv-cs-cl
4 Aug 2026
Hardware

Bole: Efficient Tree Speculation for Hybrid-Attention Language Models

DGX agent

arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autore

hardwarearxiv-cs-cl
4 Aug 2026
Model Releases

Certifying Plans under Model Mismatch: A Trilemma for Reachability from Scarce Data

DGX agent

arXiv:2608.02453v1 Announce Type: new Abstract: Sim-to-real policies are designed under nominal dynamics, but target-system trials may yield only a few isolated one-step transitions. We study pre-exec

model-releasesarxiv-cs-ro
4 Aug 2026
Model Releases

Cluster-Aware Over-the-Air Federated Learning with Energy-Harvesting Devices: From Global Training to Model Personalization

DGX agent

arXiv:2608.01426v1 Announce Type: new Abstract: Federated learning (FL) enables distributed optimization and learning across decentralized edge devices while preserving data privacy, but its performan

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models

DGX agent

arXiv:2608.01644v1 Announce Type: new Abstract: In video understanding, vision-language models (VLMs) must ingest massive numbers of visual tokens, causing the computational and memory cost of the pre

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

I benchmarked the 4 models I had pulled. The 1.1GB one beat the 2GB one at math and lost badly at extraction.

DGX agent

152 generations, deterministic grading (exact number/string/JSON/regex), no LLM judge on my 16GB laptop. task type | deepseek-r1:1.5b (1.1GB) | llama3.2:3b (2.0GB) | gemma:2b | codellama 7b arithmetic

model-releasesr-ollama
4 Aug 2026
Applications

Learning the Pareto Frontier of Predictive Models under Distribution Shift

DGX agent

arXiv:2608.00632v1 Announce Type: new Abstract: Modern machine learning pipelines increasingly rely on reusing pretrained and foundation models across downstream tasks. These pretrained models can dif

applicationsarxiv-cs-lg
4 Aug 2026
Model Releases

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models

DGX agent

arXiv:2608.02197v1 Announce Type: new Abstract: Visual representations of VLA models remain unreliable for spatially precise robotic manipulation. We uncover that vision encoders in VLAs also exhibit

model-releasesarxiv-cs-ro
4 Aug 2026
Safety

MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models

DGX agent

arXiv:2608.01012v1 Announce Type: new Abstract: Uncommon and off-guideline cases are difficult for clinical decision support, because physicians must make a series of management decisions under diagno

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 (Mistral AI Blog)

DGX agent

Mistral AI Blog: Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 — Every product that ships a m

model-releasestechmeme
4 Aug 2026
Safety

MMPhysVideo: Physically Plausible Video Generation Through Joint RGB-Perception Modeling

DGX agent

arXiv:2604.02817v2 Announce Type: replace Abstract: Despite advancements in generating visually stunning content, video diffusion models (VDMs) often yield physically inconsistent results due to pixel

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

Obshazard-bench: Benchmarking Multimodal Foundation Models for Real-Time Disaster Intelligence from Raw Earth Observation Streams

DGX agent

arXiv:2608.00012v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly used to interpret Earth observation data, yet their capability to support real-world disaster

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

PALMs: Using Multi Construct-Grounded Rationales for Modeling Population Preferences in LLMs

DGX agent

arXiv:2608.01458v1 Announce Type: new Abstract: Large language models are being extensively used to simulate individual user behavior, yet faithfully representing a population requires capturing the s

safetyarxiv-cs-cl
4 Aug 2026
Research

Progressive^2: A Teacher-Student Progressive Co-Evolving Knowledge Distillation Method for Substantial Model Compression

DGX agent

arXiv:2608.00129v1 Announce Type: new Abstract: Knowledge distillation (KD) is a widely utilized technique for transferring knowledge from a large model (the teacher) to a smaller model (the student).

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from…

DGX agent

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from RTX 4090 or equivalent). New proprietary non-decoder-only a

model-releasesyohei-nakajima--x
4 Aug 2026
Model Releases

Role Steering of Language Models for Social Simulations

DGX agent

arXiv:2608.00023v1 Announce Type: new Abstract: Social simulations built from language-model agents need role-conditioned behavior that can be checked before agents are placed into a simulated populat

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Self-Improving Large Language Models via Progressive Experience Evolution

DGX agent

arXiv:2608.02139v1 Announce Type: new Abstract: Large language models (LLMs) capable of self-improvement require not only effective policy optimization, but also a principled mechanism for transformin

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

A Unified Benchmark of Deep Learning Models for Multi-task 3D Brain Tumor Segmentation from Magnetic Resonance Imaging

DGX agent

arXiv:2607.28858v1 Announce Type: cross Abstract: Automatic brain tumor segmentation from magnetic resonance imaging (MRI) has become a fundamental task in computer-assisted diagnosis, treatment plann

model-releasesarxiv-cs-ai
3 Aug 2026
← Previous
1…979899100101…1261
Next →