AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,521 results
Model Releases

🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, and Qwen3.6-27B punches way above i…

DGX agent

🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, and Qwen3.6-27B punches way above its weight. 👇 What's new: 🧠 Outstanding agentic coding — surpa

model-releasesjeremy-howard--x
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Not the first time either - they shut down a bunch of of their original proprietary hosted embedding models in this announcement back in Apr…

DGX agent

Not the first time either - they shut down a bunch of of their original proprietary hosted embedding models in this announcement back in April 2024 https://openai.com/index/gpt-4-api-general-availabil

model-releasessimon-willison--x
22 Apr 2026
Industry

OpenAI releases Privacy Filter, an open-weight model for masking personally identifiable information in text, with 1.5B total and 50M active parameters (OpenAI)

DGX agent

OpenAI: OpenAI releases Privacy Filter, an open-weight model for masking personally identifiable information in text, with 1.5B total and 50M active parameters — Our state of the art model for masking

industrytechmeme
22 Apr 2026
Research

Pause or Fabricate? Training Language Models for Grounded Reasoning

DGX agent

arXiv:2604.19656v1 Announce Type: new Abstract: Large language models have achieved remarkable progress on complex reasoning tasks. However, they often implicitly fabricate information when inputs are

researcharxiv-cs-cl
22 Apr 2026
Industry

Qwen3.6-35B-A3B is trending at #1 on Hugging Face! 🥇🤗 Thank you for making us the top trending model on @huggingface this week. Let's keep…

DGX agent

Qwen3.6-35B-A3B has achieved #1 trending status on Hugging Face, indicating significant community interest and adoption of this large language model. The announcement highlights the model's popularity

industryclem-delangue--x
22 Apr 2026
Model Releases

RepIt: Steering Language Models with Concept-Specific Refusal Vectors

DGX agent

arXiv:2509.13281v5 Announce Type: replace Abstract: Current safety evaluations of language models rely on benchmark-based assessments that may miss localized vulnerabilities. We present RepIt, a simpl

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling

DGX agent

arXiv:2604.19734v1 Announce Type: cross Abstract: Scaling humanoid foundation models is bottlenecked by the scarcity of robotic data. While massive egocentric human data offers a scalable alternative,

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

VLA Foundry: A Unified Framework for Training Vision-Language-Action Models

DGX agent

arXiv:2604.19728v1 Announce Type: cross Abstract: We present VLA Foundry, an open-source framework that unifies LLM, VLM, and VLA training in a single codebase. Most open-source VLA efforts specialize

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17

DGX agent

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17 People are misreading the SpaceX/Cursor deal as an M&A story. It’s actually a b

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection

DGX agent

arXiv:2503.11838v2 Announce Type: replace Abstract: Sarcasm detection, with its figurative nature, poses unique challenges for affective systems designed to perform sentiment analysis. While these sys

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis

DGX agent

arXiv:2604.16729v1 Announce Type: new Abstract: State-of-the-art large language models (LLMs) show high performance in general visual question answering. However, a fundamental limitation remains: cur

model-releasesarxiv-cs-cv
21 Apr 2026
Local Ai

Aligning Language Models with Real-time Knowledge Editing

DGX agent

arXiv:2508.01302v3 Announce Type: replace Abstract: Knowledge editing aims to modify outdated knowledge in language models efficiently while retaining their original capabilities. Mainstream datasets

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

AnchorMem: Anchored Facts with Associative Contexts for Building Memory in Large Language Models

DGX agent

arXiv:2604.17377v1 Announce Type: new Abstract: While large language models have achieved remarkable performance in complex tasks, they still need a memory system to utilize historical experience in l

model-releasesarxiv-cs-cl
21 Apr 2026
Industry

Anthropic's Mythos has been accessed by a small group of unauthorized users, raising questions about control of the AI model https://www.blo…

DGX agent

Anthropic's Mythos has been accessed by a small group of unauthorized users, raising questions about control of the AI model https://www.bloomberg.com/news/articles/2026-04-21/anthropic-s-mythos-model

industryclem-delangue--x
21 Apr 2026
Research

BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation

DGX agent

arXiv:2604.16514v1 Announce Type: new Abstract: Autoregressive vision-language models (VLMs) deliver strong multimodal capability, but their token-by-token decoding imposes a fundamental inference bot

researcharxiv-cs-cv
21 Apr 2026
Model Releases

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models

DGX agent

arXiv:2511.16857v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have achieved impressive performance on spatial reasoning benchmarks, yet these evaluations mask critical weaknesses i

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Can Large Language Models Understand Context?

DGX agent

Understanding context is key to understanding human language, an ability which Large Language Models (LLMs) have been increasingly seen to demonstrate to an impressive extent. However, though the eval

model-releasesapple-ml-research
21 Apr 2026
Tutorials

ControlAudio: Tackling Text-Guided, Timing-Indicated and Intelligible Audio Generation via Progressive Diffusion Modeling

DGX agent

arXiv:2510.08878v3 Announce Type: replace-cross Abstract: Text-to-audio (TTA) generation with fine-grained control signals, e.g., precise timing control or intelligible speech content, has been explor

tutorialsarxiv-cs-cl
21 Apr 2026
Research

Emergent Structured Representations Support Flexible In-Context Inference in Large Language Models

DGX agent

arXiv:2602.07794v3 Announce Type: replace Abstract: Large language models (LLMs) exhibit emergent behaviors suggestive of human-like reasoning. While recent work has identified structured conceptual r

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Enhancing Continual Learning of Vision-Language Models via Dynamic Prefix Weighting

DGX agent

arXiv:2604.18075v1 Announce Type: new Abstract: We investigate recently introduced domain-class incremental learning scenarios for vision-language models (VLMs). Recent works address this challenge us

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Evalet: Evaluating Large Language Models through Functional Fragmentation

DGX agent

arXiv:2509.11206v4 Announce Type: replace-cross Abstract: Practitioners increasingly rely on Large Language Models (LLMs) to evaluate generative AI outputs through 'LLM-as-a-Judge' approaches. However

researcharxiv-cs-cl
21 Apr 2026
Research

Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality

DGX agent

arXiv:2505.11140v3 Announce Type: replace Abstract: We introduce fs1, a simple yet effective method that improves the factuality of reasoning traces by collecting them from large reasoning models and

researcharxiv-cs-cl
21 Apr 2026
Applications

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models

DGX agent

arXiv:2603.04592v3 Announce Type: replace Abstract: Standard Large Language Models (LLMs) are predominantly designed for static inference with pre-defined inputs, which limits their applicability in d

applicationsarxiv-cs-cl
21 Apr 2026
Research

HiPrune: Hierarchical Attention for Efficient Token Pruning in Vision-Language Models

DGX agent

arXiv:2508.00553v3 Announce Type: replace Abstract: Vision-Language Models (VLMs) encode images and videos into abundant tokens, which contain substantial redundancy and computation cost. While visual

researcharxiv-cs-cv
21 Apr 2026
Model Releases

HQA-VLAttack: Towards High Quality Adversarial Attack on Vision-Language Pre-Trained Models

DGX agent

arXiv:2604.16499v1 Announce Type: new Abstract: Black-box adversarial attack on vision-language pre-trained models is a practical and challenging task, as text and image perturbations need to be consi

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions

DGX agent

arXiv:2601.05414v2 Announce Type: replace Abstract: As large language models (LLMs) transition from chat interfaces to integral components of stochastic pipelines and systems approaching general intel

applicationsarxiv-cs-cl
21 Apr 2026
Research

Linear-Time and Constant-Memory Text Embeddings Based on Recurrent Language Models

DGX agent

arXiv:2604.18199v1 Announce Type: new Abstract: Transformer-based embedding models suffer from quadratic computational and linear memory complexity, limiting their utility for long sequences. We propo

researcharxiv-cs-cl
21 Apr 2026
Safety

Mammo-FM: Breast-specific foundational model for Integrated Mammographic Diagnosis, Prognosis, and Reporting

DGX agent

arXiv:2512.00198v2 Announce Type: replace Abstract: Breast cancer is one of the leading causes of death among women worldwide. We introduce Mammo-FM, the first foundation model specifically for mammog

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

DGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

DGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage

DGX agent

arXiv:2604.17354v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at photorealistic generation, yet often struggle to represent abstract meaning such as idiomatic interpretations of

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

OneDrive: Unified Multi-Paradigm Driving with Vision-Language-Action Models

DGX agent

arXiv:2604.17915v1 Announce Type: new Abstract: Vision-Language Models(VLMs) excel at autoregressive text generation, yet end-to-end autonomous driving requires multi-task learning with structured out

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

Prompting Foundation Models for Zero-Shot Ship Instance Segmentation in SAR Imagery

DGX agent

arXiv:2604.17920v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) plays a critical role in maritime surveillance, yet deep learning for SAR analysis is limited by the lack of pixel-level

model-releasesarxiv-cs-cv
21 Apr 2026
Tutorials

RePrompT: Recurrent Prompt Tuning for Integrating Structured EHR Encoders with Large Language Models

DGX agent

arXiv:2604.17725v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong promise for mining Electronic Health Records (EHRs) by reasoning over longitudinal clinical information t

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models

DGX agent

arXiv:2604.17691v1 Announce Type: new Abstract: Safety alignment in large language models is remarkably shallow: it is concentrated in the first few output tokens and reversible by fine-tuning on as f

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

SFTMix: Elevating Language Model Instruction Tuning with Mixup Recipe

DGX agent

arXiv:2410.05248v4 Announce Type: replace Abstract: To acquire instruction-following capabilities, large language models (LLMs) undergo instruction tuning, where they are trained on instruction-respon

applicationsarxiv-cs-cl
21 Apr 2026
Research

Stable Language Guidance for Vision-Language-Action Models

DGX agent

arXiv:2601.04052v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive capabilities in generalized robotic control; however, they remain notoriously

researcharxiv-cs-cl
21 Apr 2026
Local Ai

Topology-Aware Layer Pruning for Large Vision-Language Models

DGX agent

arXiv:2604.16502v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in natural language understanding and reasoning, while recent extensions that incorpo

local-aiarxiv-cs-cv
21 Apr 2026
Safety

Toward Consistent World Models with Multi-Token Prediction and Latent Semantic Enhancement

DGX agent

arXiv:2604.06155v2 Announce Type: replace-cross Abstract: Whether Large Language Models (LLMs) develop coherent internal world models remains a core debate. While conventional Next-Token Prediction (N

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

UniMamba: A Unified Spatial-Temporal Modeling Framework with State-Space and Attention Integration

DGX agent

arXiv:2604.16325v1 Announce Type: new Abstract: Multivariate time series forecasting is fundamental to numerous domains such as energy, finance, and environmental monitoring, where complex temporal de

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Unleashing Spatial Reasoning in Multimodal Large Language Models via Textual Representation Guided Reasoning

DGX agent

arXiv:2603.23404v2 Announce Type: replace-cross Abstract: Existing Multimodal Large Language Models (MLLMs) struggle with 3D spatial reasoning, as they fail to construct structured abstractions of the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2604.18000v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models report impressive success rates on standard robotic benchmarks, fueling optimism about general-purpose physic

model-releasesarxiv-cs-ro
21 Apr 2026
Research

Aligning What Vision-Language Models See and Perceive with Adaptive Information Flow

DGX agent

arXiv:2604.15809v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated strong capability in a wide range of tasks such as visual recognition, document parsing, and visual grou

researcharxiv-cs-cv
20 Apr 2026
Model Releases

DALM: A Domain-Algebraic Language Model via Three-Phase Structured Generation

DGX agent

arXiv:2604.15593v1 Announce Type: cross Abstract: Large language models compress heterogeneous knowledge into a single parameter space, allowing facts from different domains to interfere during genera

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3…

DGX agent

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3.6-Plus 📖 Stronger world knowledge and instruction following

model-releasesqwen--x
20 Apr 2026
Research

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models

DGX agent

arXiv:2604.15741v1 Announce Type: cross Abstract: Uncertainty estimation is a promising approach to detect hallucinations in large language models (LLMs). Recent approaches commonly depend on model in

researcharxiv-cs-ai
20 Apr 2026
Research

Noise Aggregation Analysis Driven by Small-Noise Injection: Efficient Membership Inference for Diffusion Models

DGX agent

arXiv:2510.21783v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated powerful performance in generating high-quality images. A typical example is text-to-image generator like S

researcharxiv-cs-ai
20 Apr 2026
Model Releases

P3T: Prototypical Point-level Prompt Tuning with Enhanced Generalization for 3D Vision-Language Models

DGX agent

arXiv:2604.15703v1 Announce Type: new Abstract: With the rise of pre-trained models in the 3D point cloud domain for a wide range of real-world applications, adapting them to downstream tasks has beco

model-releasesarxiv-cs-cv
20 Apr 2026
← Previous
1…117118119120121…1261
Next →