AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,536 results
Research

On-board Remote-Sensing Foundation Models for Unsupervised Change Detection of Disaster Events

DGX agent

arXiv:2606.27018v1 Announce Type: cross Abstract: Remote Sensing Foundation Models (RSFMs) have emerged as a powerful alternative to supervised models for Earth Observation, allowing satellites to aut

researcharxiv-cs-ai
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

PhysEditWorld: A Large-Scale Dataset Toward Physics-Editable World Models

DGX agent

arXiv:2606.26694v1 Announce Type: new Abstract: Recent game world models can synthesize visually plausible, action-conditioned rollouts. However, their interaction behaviors often remain limited to ex

safetyarxiv-cs-cv
26 Jun 2026
Industry

There's a lot of sloppy thinking around open models. You can ban them and make it impossible for US companies to use them, but this won't st…

DGX agent

There's a lot of sloppy thinking around open models. You can ban them and make it impossible for US companies to use them, but this won't stop A) global open model progress B) bad actors using them So

industryclem-delangue--x
26 Jun 2026
Model Releases

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

DGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

model-releasesallie-k--miller--x
26 Jun 2026
Tutorials

Transformers are better at copying, while RNNs are better at modeling 'meaning-bearing words—the nouns, verbs, & adjectives that say what a …

DGX agent

Transformers are better at copying, while RNNs are better at modeling 'meaning-bearing words—the nouns, verbs, & adjectives that say what a sentence is about' Hybrid (transformer–RNN) models are fast

tutorialsjeremy-howard--x
26 Jun 2026
Model Releases

Tuning Language Models by Mixture-of-Depths Ensemble

DGX agent

arXiv:2410.13077v2 Announce Type: replace-cross Abstract: Transformer-based Large Language Models (LLMs) traditionally rely on final-layer loss for finetuning and final-layer representations for predi

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. …

DGX agent

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. Faulty reasoning, use of inappropriate shortcuts, hallucinat

model-releasesgary-marcus--x
26 Jun 2026
Model Releases

Agentic evolution of physically constrained foundation models

DGX agent

arXiv:2606.25532v1 Announce Type: cross Abstract: Artificial intelligence increasingly drives automated scientific discovery, yet contemporary generalist agents lack physical grounding, frequently hal

model-releasesarxiv-cs-lg
25 Jun 2026
Research

Conformal Orbit-Valid Trust Horizons for Equivariant World Models

DGX agent

arXiv:2606.24946v1 Announce Type: new Abstract: Learned world models are useful only over horizons on which their rollout error remains controlled. We study trust-horizon certification for latent worl

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Elo-Disentangled Player-Style Embeddings for Human Chess via Rating-Conditioned Residual Move Model

DGX agent

arXiv:2606.25176v1 Announce Type: new Abstract: We study representation learning for individual human chess style: a per-player embedding learned from a player's move history such that inner products

model-releasesarxiv-cs-ai
25 Jun 2026
Industry

Feds deny Polestar authorization to sell cars in US from model year 2027

DGX agent

The U.S. has denied authorization for Polestar to sell 2027 model-year vehicles, effectively preventing new Polestar models from entering the U.S. market. The Connected Vehicle Rule restricts vehicles

industryars-technica
25 Jun 2026
Model Releases

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

DGX agent

arXiv:2606.25391v1 Announce Type: cross Abstract: Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including sp

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣

DGX agent

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣 llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance j

model-releasesgeorgi-gerganov--x
25 Jun 2026
Model Releases

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models

DGX agent

arXiv:2606.25402v1 Announce Type: cross Abstract: Large software projects often depend on older versions of libraries, even as APIs continue to evolve across releases. This creates a challenge for LLM

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources

DGX agent

arXiv:2606.25832v1 Announce Type: new Abstract: Achieving strong optimization generalization across diverse optimization problems while requiring limited training resources remains a challenging probl

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

Privacy-Aware Visual Language Models

DGX agent

arXiv:2405.17423v4 Announce Type: replace-cross Abstract: As Visual Language Models (VLMs) become increasingly embedded in everyday applications, ensuring they can recognise and appropriately handle p

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model…

DGX agent

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model in one place. We price by GPU-hour, not per token, so long

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversational Models

DGX agent

arXiv:2606.25990v1 Announce Type: new Abstract: As multimodal conversational systems increasingly engage in spoken interaction, their ability to navigate paralinguistic social cues has become a critic

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Wan-Streamer v0.1: End-to-end Real-time Interactive Foundation Models

DGX agent

arXiv:2606.25041v1 Announce Type: new Abstract: We present Wan-Streamer, a native-streaming, end-to-end interactive foundation model designed from the ground up for real-time, low-latency, full-duplex

researcharxiv-cs-cv
25 Jun 2026
Model Releases

Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War

DGX agent

arXiv:2606.24391v1 Announce Type: new Abstract: We introduce Age of LLM, a turn-based 1v1 benchmark in which two LLMs face off on a 13x7 grid to destroy the enemy base. Three stressors are deliberate:

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

An analysis of GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Gab's Arya, and other AI models: most chatbots frequently provide left-leaning responses to political prompts (Kevin Schaul/Washington Post)

DGX agent

Kevin Schaul / Washington Post: An analysis of GPT-5.5, Gemini 3.1 Pro, Grok 4.3, Gab's Arya, and other AI models: most chatbots frequently provide left-leaning responses to political prompts — Excerp

model-releasestechmeme
24 Jun 2026
Agents

Autonomous Video Generation with Counterfactual Controllability for Self-Evolving World Models

DGX agent

arXiv:2606.24152v1 Announce Type: new Abstract: Existing literature claims that video generation essentially is world modelling. On the one hand, the claim is productive because it pushes generative A

agentsarxiv-cs-cv
24 Jun 2026
Safety

CALIBER: Calibrating Confidence Before and After Reasoning in Language Models

DGX agent

arXiv:2606.24281v1 Announce Type: cross Abstract: Reasoning language models are increasingly asked not only to answer difficult questions, but also to estimate their likelihood of success. Existing me

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

FISHER: A Foundation Model for Multi-Modal Industrial Signal Comprehensive Representation

DGX agent

arXiv:2507.16696v3 Announce Type: replace-cross Abstract: Industrial signal analysis is hindered by severe data heterogeneity, which we characterize as the M5 problem. Existing solutions rely on speci

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Geometric Action Model for Robot Policy Learning

DGX agent

arXiv:2606.17046v2 Announce Type: replace-cross Abstract: Generalist robot policies must follow user instructions while reasoning about how objects, cameras, and robot actions interact in the 3D physi

safetyarxiv-cs-cv
24 Jun 2026
Tools

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest …

DGX agent

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest open-source frontier model without changing your workflow. Y

toolsfireworks-ai--x
24 Jun 2026
Model Releases

L3Cube-MahaPOS: A Marathi Part-of-Speech Tagging Dataset and BERT Models

DGX agent

arXiv:2606.24825v1 Announce Type: new Abstract: Part-of-Speech (POS) tagging is a foundational NLP task underpinning machine translation, information extraction, and syntactic parsing. Despite Marathi

model-releasesarxiv-cs-cl
24 Jun 2026
Research

Maestro Order: A Model-Agnostic Orchestration Harness

DGX agent

arXiv:2606.23983v1 Announce Type: cross Abstract: A single forward pass of a capable model is a fast, fluent, and unreliable problem-solver: it is right often enough to be useful and wrong often enoug

researcharxiv-cs-ai
24 Jun 2026
Model Releases

MambaRaw: Selective State Space Modeling for Efficient 4K Raw Image Reconstruction

DGX agent

arXiv:2606.24479v1 Announce Type: new Abstract: In-camera JPEG previews are ubiquitous in raw image formats and provide an sRGB reference at negligible storage cost. Although existing metadata-based r

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

MedBench v5: A Dynamic, Process-Oriented, and Hallucination-Aware Benchmark for Clinical Multimodal Models

DGX agent

arXiv:2606.24155v1 Announce Type: new Abstract: Existing medical AI benchmarks lack process visibility, atomic skill evaluation, and integrated hallucination detection. We introduce MedBench v5, a red

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

On the Stability of Prompt Ranking in Large Language Model Evaluation

DGX agent

arXiv:2606.24381v1 Announce Type: cross Abstract: Prompt-based interaction has become a dominant paradigm for using large language models (LLMs), where multiple candidate prompts are evaluated and the

model-releasesarxiv-cs-ai
24 Jun 2026
Research

Pigeonholing: Bad prompts hurt models to collapse and make mistakes

DGX agent

arXiv:2606.24267v1 Announce Type: cross Abstract: While in-context learning is generally shown to be effective in Large Language Models (LLMs), bad contexts can cause performance degradation and mode

researcharxiv-cs-ai
24 Jun 2026
Safety

Reinforcement Learning Towards Broadly and Persistently Beneficial Models

DGX agent

arXiv:2606.24014v1 Announce Type: new Abstract: As AI systems are deployed across increasingly diverse and high-stakes settings, model alignment must generalize beyond the tasks and domains seen durin

safetyarxiv-cs-ai
24 Jun 2026
Research

S1-Omni-Image: A Unified Model for Scientific Image Understanding, Generation, and Editing

DGX agent

arXiv:2606.24441v1 Announce Type: new Abstract: We present S1-Omni-Image, an open-weight unified multimodal model for scientific image understanding, generation, and editing. Unlike general-purpose im

researcharxiv-cs-cv
24 Jun 2026
Safety

ScaleToT: Generalizing Structured LLM Reasoning for Billion-Scale Low-Activity User Modeling

DGX agent

arXiv:2606.24605v1 Announce Type: new Abstract: Accurate user modeling often depends on rich interaction histories, which are unavailable for billions of low-activity users. Large Language Models (LLM

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

This is the strongest ARC-AGI-2 performance to date by an open-source model.

DGX agent

This is the strongest ARC-AGI-2 performance to date by an open-source model. GLM-5.2 from @Zai_org on ARC-AGI (Verified) - ARC-AGI-2: 22.8%, 0.25 - ARC-AGI-1: 77.0%, 0.19 Performance is comparable wit

model-releasesfrancois-chollet--x
24 Jun 2026
Model Releases

We have a new version of GPT-5.5 Instant for you, and it's much more fun to talk to. Our most-used model is now better at understanding the …

DGX agent

We have a new version of GPT-5.5 Instant for you, and it's much more fun to talk to. Our most-used model is now better at understanding the intent behind a question and adapting its response according

model-releasesopenai--x
24 Jun 2026
Safety

When Preferences Fail to Become Incentives: A Utility-Behavior Gap in Large Language Models

DGX agent

arXiv:2606.22974v2 Announce Type: replace Abstract: Recent work on preference elicitation in large language models (LLMs) has demonstrated that, when given a series of choices between two outcomes, LL

safetyarxiv-cs-ai
24 Jun 2026
Safety

BadDreamer: Transferable Backdoor Attacks against Video World Models for Autonomous Driving

DGX agent

arXiv:2606.21172v1 Announce Type: new Abstract: Video world models are increasingly used in autonomous driving to forecast future scene evolution and provide future-aware spatio-temporal representatio

safetyarxiv-cs-cv
23 Jun 2026
Research

Certified World Models: Predictability Across Configuration, Horizon, and Resolution

DGX agent

arXiv:2606.13092v2 Announce Type: replace Abstract: Scale buys interpolation; structure buys certifiable transfer. A world model's average error does not say whether a particular rollout can be truste

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Detail++: Training-Free Detail Enhancer for T2I Diffusion Models

DGX agent

arXiv:2507.17853v3 Announce Type: replace Abstract: Recent advances in text-to-image (T2I) generation have led to impressive visual results. However, these models still face significant challenges whe

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Does RoPE Prevent or Degrade Retrieval Heads? A Mechanistic Analysis Across Model Families

DGX agent

arXiv:2606.21249v1 Announce Type: new Abstract: Retrieval heads, attention heads that copy information from earlier context to the current position, have been proposed as the mechanistic substrate for

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

dVLA-RL: Reinforcement Learning over Denoising Trajectories for Discrete Diffusion Vision-Language-Action Models

DGX agent

arXiv:2606.23623v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have established a powerful paradigm for generalist robotic manipulation by grounding control into the semantic reas

safetyarxiv-cs-ro
23 Jun 2026
Safety

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning

DGX agent

arXiv:2606.22860v1 Announce Type: new Abstract: Recent advancements in generative imitation learning have significantly propelled the field of robotic manipulation. However, the majority of existing m

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

Large Language Model-Assisted Cleaning of Report-Derived Labels in a Large-Scale Chest CT Dataset

DGX agent

arXiv:2606.22382v1 Announce Type: cross Abstract: Purpose: To evaluate whether large language model (LLM)-assisted label cleaning can identify label-report discordance in CT-RATE, a large-scale public

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Mistral debuts OCR 4, a model featuring structured document extraction with bounding boxes, block classification, and inline confidence scores, in 170 languages (Mistral AI Blog)

DGX agent

Mistral AI Blog: Mistral debuts OCR 4, a model featuring structured document extraction with bounding boxes, block classification, and inline confidence scores, in 170 languages — Today, we're releasi

model-releasestechmeme
23 Jun 2026
Research

NeuroShield: A Device-Agnostic Foundation Model for EEG Authentication

DGX agent

arXiv:2606.20673v1 Announce Type: cross Abstract: A central challenge in EEG authentication is that models are typically tied to the acquisition settings in which they are trained. In particular, vari

researcharxiv-cs-cv
23 Jun 2026
Model Releases

OGD4All: A Framework for Accessible Interaction with Geospatial Open Government Data Based on Large Language Models

DGX agent

arXiv:2602.00012v3 Announce Type: replace Abstract: We present OGD4All, a transparent, auditable, and reproducible framework based on Large Language Models (LLMs) to enhance citizens' interaction with

model-releasesarxiv-cs-lg
23 Jun 2026
← Previous
1…131132133134135…1262
Next →