AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,928 results
Local Ai

For the last few months, the conversation around open models has centered on cost and performance optimization. But Satya highlights somethi…

DGX agent

For the last few months, the conversation around open models has centered on cost and performance optimization. But Satya highlights something more existential: open models aren't just an optimization

local-aiollama--x
13 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents

DGX agent

arXiv:2510.22052v2 Announce Type: replace Abstract: The field of artificial intelligence (AI) has taken a tight hold on broad aspects of society, industry, business, and governance in ways that dictat

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

DeltaV: Thinking with Visual State Updates in Unified Large Multimodal Models

DGX agent

arXiv:2607.08434v1 Announce Type: new Abstract: Current Unified Large Multimodal Models (ULMMs) support interleaved multimodal reasoning through textual reasoning and intermediate visual states, but t

model-releasesarxiv-cs-cv
10 Jul 2026
Model Releases

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

DGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

DGX agent

arXiv:2607.08540v1 Announce Type: cross Abstract: Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to to

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

DGX agent

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al.

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models

DGX agent

arXiv:2607.08059v1 Announce Type: cross Abstract: Uncertainty quantification for visual language models (VLMs) conventionally targets the answer token distribution. We provide the first three-family e

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Co-LMLM: Continuous-Query Limited Memory Language Models

DGX agent

arXiv:2607.07707v1 Announce Type: cross Abstract: Limited memory language models (LMLMs) externalize factual knowledge during pretraining to a knowledge base (KB), rather than memorizing it in their w

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Does Bielik Know What It Doesn't Know? Activation Dispersion Separates Entity Familiarity from Factual Reliability Across Model Scale

DGX agent

arXiv:2607.07670v1 Announce Type: new Abstract: Large language models hallucinate most about entities they have never seen. We ask whether a model's activations betray entity familiarity before a sing

researcharxiv-cs-cl
9 Jul 2026
Model Releases

FMMC: Harnessing the Power of Foundation Models for Accurate Material Classification

DGX agent

arXiv:2603.17390v2 Announce Type: replace Abstract: Material classification has emerged as a critical task in computer vision and graphics, supporting the assignment of accurate material properties to

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

MedPMC: A Systematic Framework for Scaling High-Fidelity Medical Multimodal Data for Foundation Models

DGX agent

arXiv:2607.07673v1 Announce Type: new Abstract: Medicine is inherently multimodal, requiring clinicians to synthesize information across diverse data streams. Yet the development of multimodal foundat

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Muse Spark 1.1 available in new Meta Model API. Somewhere near Opus-4.8/GPT-5.5 level. 1M context window! Computer-use capabilities sound gr…

DGX agent

Muse Spark 1.1 available in new Meta Model API. Somewhere near Opus-4.8/GPT-5.5 level. 1M context window! Computer-use capabilities sound great: Write scripts when automation is faster, click when dir

model-releasesdair-ai--x
9 Jul 2026
Model Releases

Sol, Terra, and Luna, our GPT‑5.6 family of models, are starting to roll out now in ChatGPT, Codex, and the API.

DGX agent

OpenAI has announced the rollout of its GPT-5.6 family of models, consisting of Sol, Terra, and Luna variants, across ChatGPT, Codex, and API platforms. These models represent the latest iteration in

model-releasesopenai--x
9 Jul 2026
Model Releases

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context wind…

DGX agent

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context window is expected to jump to 1M tokens, possibly as soon as nex

model-releaseselon-musk--x
9 Jul 2026
Model Releases

A good voice model should be enjoyable to talk to, and GPT-Live is a great conversationalist with a more natural and defined personality tha…

DGX agent

GPT-Live is OpenAI's voice model designed to be an engaging conversational partner with natural speech and a distinct personality. The model prioritizes making interactions enjoyable for users through

model-releasesopenai--x
8 Jul 2026
Model Releases

AirflowAttack: Thermal-Airflow Adversarial Perturbations against Infrared Remote-Sensing Vision-Language Models

DGX agent

arXiv:2607.06485v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on infrared (IR) remote sensing imagery in security-critical settings, yet their adversarial r

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

From Foundation to Application: Improving VLA Models in Practice

DGX agent

arXiv:2607.06403v1 Announce Type: new Abstract: Despite recent progress of VLA foundation models, the disparity between laboratory conditions and real-world applications continues to impede their prac

model-releasesarxiv-cs-ro
8 Jul 2026
Model Releases

GPT-Live makes talking with AI feel like having a real conversation. It’s also our smartest voice model yet. https://openai.com/index/introd…

DGX agent

OpenAI introduced GPT-Live, a voice-based AI model designed to provide conversational interactions that feel more natural and human-like than previous versions. The model represents an advancement in

model-releasesopenai--x
8 Jul 2026
Model Releases

Hierarchical Acoustic-Semantic Modeling: Modality Separation and Semantic Coherence for Full-Duplex SLMs

DGX agent

arXiv:2607.06540v1 Announce Type: new Abstract: Developing seamless, high-performance, native intelligent full-duplex Spoken Language Models (SLMs) remains a critical challenge and long-standing goal

model-releasesarxiv-cs-cl
8 Jul 2026
Model Releases

Introducing GPT-Live, a new generation of voice models for natural human-AI interaction. Rolling out in ChatGPT starting today. You’ll want …

DGX agent

GPT-Live is a new generation of voice model developed by OpenAI designed to enable more natural human-AI voice interactions. The model began rolling out to ChatGPT users starting on the date of this a

model-releasesopenai--x
8 Jul 2026
Model Releases

OpenAI launches GPT-Live, a new generation of voice models built on a full-duplex architecture, meaning they can listen and speak at the same time (OpenAI)

DGX agent

OpenAI: OpenAI launches GPT-Live, a new generation of voice models built on a full-duplex architecture, meaning they can listen and speak at the same time — A new generation of voice models for natura

model-releasestechmeme
8 Jul 2026
Model Releases

Refiant goes where rivals only promised with a 10 million-token AI model

DGX agent

Artificial intelligence optimization startup Refiant Inc. today launched Protea, a suite of long-context AI models led by a 10 million-token context window that the company says ranks among the larges

model-releasessiliconangle
8 Jul 2026
Model Releases

VisCoP: Visual Probing for Video Domain Adaptation of Vision Language Models

DGX agent

arXiv:2510.13808v2 Announce Type: replace Abstract: Large Vision Language Models (VLMs) excel at general visual reasoning but experience significant performance degradation when deployed in novel doma

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

Alibaba's Qwen models have made it an AI powerhouse, but the company has struggled to turn their global popularity into a profitable business (New York Times)

DGX agent

New York Times: Alibaba's Qwen models have made it an AI powerhouse, but the company has struggled to turn their global popularity into a profitable business — The Chinese company's models have won ov

model-releasestechmeme
7 Jul 2026
Model Releases

Dashboard2Code: Evaluating Multimodal Models on Reconstructing Interactive Dashboards

DGX agent

arXiv:2607.04727v1 Announce Type: cross Abstract: Automatic data visualization generation has advanced rapidly with multi-modal large language models, yet existing efforts largely focus on static char

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Do Diabetic Foot Ulcer Segmentation Models Generalize? A Cross-Dataset Benchmark of CNN and Transformer Architectures

DGX agent

arXiv:2607.02555v1 Announce Type: new Abstract: Deep learning models for diabetic foot ulcer (DFU) segmentation routinely report high accuracy, but they are almost always trained and tested on the sam

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

LILAC: Layer-Wise Independent LoRAs and Cascaded Conditioning for Multi-Concept Customization of Diffusion Models

DGX agent

arXiv:2607.04801v1 Announce Type: new Abstract: Personalizing text-to-image diffusion models to render several specific subjects in a coherent image remains challenging: the model must preserve each s

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Metronome: Bound the Cache, Keep the Beat for Real-Time Interaction Model Serving

DGX agent

arXiv:2607.02640v1 Announce Type: cross Abstract: Real-time interaction models -- Moshi, MiniCPM-o, Qwen-Omni -- turn serving into a periodic real-time task: on every frame a session ingests streaming

model-releasesarxiv-cs-ai
7 Jul 2026
Applications

One Framework for All: Cross-Modal Membership Inference for Generative Models

DGX agent

arXiv:2607.04339v1 Announce Type: cross Abstract: Large generative models across text-to-text, text-to-image, and image-to-text modalities have been shown to pose significant privacy risks. One fundam

applicationsarxiv-cs-ai
7 Jul 2026
Model Releases

Predicting Biased Human Decision-Making with Large Language Models in Conversational Settings

DGX agent

arXiv:2601.11049v2 Announce Type: replace-cross Abstract: We examine whether large language models (LLMs) can predict biased decision-making in conversational settings, and whether their predictions c

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

RABBiT: Rapidly adaptive BOLD foundation model via brain-tuning for accurate zero-shot and few-shot prediction of speech-elicited responses in the brain

DGX agent

arXiv:2607.05171v1 Announce Type: new Abstract: Language understanding in the brain is context-dependent, varying across experimental stimuli and individuals, which makes it difficult to build computa

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Reconstruction-Anchored Diffusion Model for Text-to-Motion Generation

DGX agent

arXiv:2601.14788v2 Announce Type: replace Abstract: Diffusion models have seen widespread adoption for text-driven human motion generation and related tasks due to their impressive generative capabili

model-releasesarxiv-cs-cv
7 Jul 2026
Local Ai

TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training

DGX agent

arXiv:2607.02840v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown promising generalization in robotic manipulation, but they still struggle with contact-rich tasks, where

local-aiarxiv-cs-ro
7 Jul 2026
Model Releases

Topology-Driven Transferability Estimation for 3D Medical Vision Foundation Models

DGX agent

arXiv:2607.04199v1 Announce Type: new Abstract: The growing number of medical vision foundation models highlights the need for effective model selection. However, mainstream selection methods rely on

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Variable Bit-width Quantization: Learning Per-Group Precision for 'Bigger-but-Smaller' Language Models

DGX agent

arXiv:2607.02893v1 Announce Type: cross Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introdu

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

WSA_1: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control

DGX agent

arXiv:2607.03941v1 Announce Type: new Abstract: Recent advances in embodied AI have established robot foundation models (RFMs) as the dominant approach for generalist robotic systems to date. By lever

model-releasesarxiv-cs-ro
7 Jul 2026
Research

Revisiting ASR Error Correction with Specialized Models

DGX agent

Language models play a central role in automatic speech recognition (ASR), yet most methods rely on text-only models unaware of ASR error patterns. Recently, large language models (LLMs) have been app

researchapple-ml-research
6 Jul 2026
Model Releases

BRIDGE: Predicting Human Task Completion Time From Model Performance

DGX agent

arXiv:2602.07267v2 Announce Type: replace Abstract: Evaluating the real-world capabilities of AI systems requires grounding benchmark performance in human-interpretable measures of task difficulty. Ex

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Locality-Aware Continual Unlearning for Diffusion Models

DGX agent

arXiv:2512.02657v2 Announce Type: replace-cross Abstract: Real-world deployment of text-to-image diffusion models requires continual concept removal as new privacy, copyright, or safety obligations ar

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation

DGX agent

arXiv:2607.01938v1 Announce Type: cross Abstract: Manipulating fast and dynamically moving targets in unstructured 3D environments remains challenging for embodied AI. Existing visual-language-action

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Auditing Forgetting in Limited Memory Language Models

DGX agent

arXiv:2607.00605v1 Announce Type: cross Abstract: Limited Memory Language Models (LMLMs) externalize factual knowledge to a database to enable deletion-based unlearning without retraining. Existing ev

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

From Structural Equation Modelling to Double Machine Learning: Robustness Analysis for Survey-Based Research

DGX agent

arXiv:2607.00512v1 Announce Type: new Abstract: Structural equation modelling (SEM) is widely used in survey-based business and information systems research to assess latent constructs and theory-driv

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models

DGX agent

arXiv:2607.01117v1 Announce Type: new Abstract: Video Large Language Models (VideoLLMs) have shown strong progress in video understanding, yet they still suffer from hallucinations that are inconsiste

model-releasesarxiv-cs-cv
2 Jul 2026
Safety

Revisiting Autoregressive Models for Generative Image Classification

DGX agent

arXiv:2603.19122v2 Announce Type: replace Abstract: Class-conditional generative models have emerged as accurate and robust classifiers, with diffusion models demonstrating clear advantages over other

safetyarxiv-cs-cv
2 Jul 2026
Model Releases

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models

DGX agent

arXiv:2603.16859v2 Announce Type: replace Abstract: Omni-modal large language models (OLMs) redefine human-machine interaction by natively integrating audio, vision, and text. However, existing OLM be

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Testing Frontier Large Language Models' Physics Literacy in Parallel Physical Worlds

DGX agent

arXiv:2607.00276v1 Announce Type: cross Abstract: Current large-language-model (LLM) physics benchmarks are usually scored by answer accuracy, which cannot distinguish genuine reasoning from recall of

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

UniDrive-WM: Unified Understanding, Planning and Generation World Model for Autonomous Driving

DGX agent

arXiv:2601.04453v4 Announce Type: replace Abstract: World models have become central to autonomous driving, where accurate scene understanding and future prediction are crucial for safe control. Recen

model-releasesarxiv-cs-cv
2 Jul 2026
Research

Valdi: Value Diffusion World Models

DGX agent

arXiv:2607.00917v1 Announce Type: cross Abstract: World models can enable Model Predictive Control (MPC), but this requires dynamics prediction that is both fast enough for online use and expressive e

researcharxiv-cs-ai
2 Jul 2026
← Previous
1…5960616263…1249
Next →