AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
Model Releases

Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework

DGX agent

arXiv:2604.10535v1 Announce Type: cross Abstract: Incorporating large language models (LLMs) in medical question answering demands more than high average accuracy: a model that returns substantively d

model-releasesarxiv-cs-cl
14 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

From Perception to Autonomous Computational Modeling: A Multi-Agent Approach

DGX agent

arXiv:2604.06788v2 Announce Type: replace-cross Abstract: We present a solver-agnostic framework in which coordinated large language model (LLM) agents autonomously execute the complete computational

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

GenTac: Generative Modeling and Forecasting of Soccer Tactics

DGX agent

arXiv:2604.11786v1 Announce Type: new Abstract: Modeling open-play soccer tactics is a formidable challenge due to the stochastic, multi-agent nature of the game. Existing computational approaches typ

model-releasesarxiv-cs-ai
14 Apr 2026
Industry

Helical raises $10M to bridge the gap between foundation models and drug discovery decisions

DGX agent

Pharma artificial intelligence startup Helical Ltd. announced today that it has raised 10 million in new funding to expand its virtual AI lab platform, which turns biological foundation models into re

industrysiliconangle
14 Apr 2026
Model Releases

How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities

DGX agent

arXiv:2603.02578v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive domains, yet their unpredictable behaviors, ranging from misalign

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Intersectional Sycophancy: How Perceived User Demographics Shape False Validation in Large Language Models

DGX agent

arXiv:2604.11609v1 Announce Type: new Abstract: Large language models exhibit sycophantic tendencies--validating incorrect user beliefs to appear agreeable. We investigate whether this behavior varies

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, th…

DGX agent

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, the biggest reason why people are struggling with open/local m

agentsclem-delangue--x
14 Apr 2026
Model Releases

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude…

DGX agent

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude Code & Codex were viable. Before that, it was all about cod

model-releasesethan-mollick--x
14 Apr 2026
Model Releases

LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models

DGX agent

arXiv:2604.09712v1 Announce Type: cross Abstract: Spatial reasoning is a cornerstone capability for intelligent systems to perceive and interact with the physical world. However, multimodal large lang

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

let's go open-source and local models!

DGX agent

let's go open-source and local models! Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its 2026 AI budget 📈 “I'm back to the drawing boar

model-releasesclem-delangue--x
14 Apr 2026
Safety

Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Adoption in Cybersecurity Operations on Reddit

DGX agent

arXiv:2604.09998v1 Announce Type: cross Abstract: Large language models (LLMs) have recently emerged as promising tools for augmenting Security Operations Center (SOC) workflows, with vendors increasi

safetyarxiv-cs-ai
14 Apr 2026
Research

Locket: Robust Feature-Locking Technique for Language Models

DGX agent

arXiv:2510.12117v3 Announce Type: replace-cross Abstract: Chatbot service providers (e.g., OpenAI) rely on tiered subscription plans to generate revenue, offering black-box access to basic models for

researcharxiv-cs-lg
14 Apr 2026
Research

MIXAR: Scaling Autoregressive Pixel-based Language Models to Multiple Languages and Scripts

DGX agent

arXiv:2604.11575v1 Announce Type: new Abstract: Pixel-based language models are gaining momentum as alternatives to traditional token-based approaches, promising to circumvent tokenization challenges.

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text

DGX agent

arXiv:2604.10151v1 Announce Type: new Abstract: Large language models are increasingly used as writing tools and pedagogical resources in English for Academic Purposes, but it remains unclear whether

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Physics-informed AI Accelerated Retention Analysis of Ferroelectric Vertical NAND: From Day-Scale TCAD to Second-Scale Surrogate Model

DGX agent

arXiv:2603.06881v2 Announce Type: replace-cross Abstract: Ferroelectric field-effect transistors (FeFET)-based vertical NAND (Fe-VNAND) has emerged as a promising candidate to overcome z-scaling limit

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models

DGX agent

arXiv:2604.09824v1 Announce Type: cross Abstract: Vision language action (VLA) models enable generalist robotic agents but often exhibit language ignorance, relying on visual shortcuts and remaining i

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Seg2Change: Adapting Open-Vocabulary Semantic Segmentation Model for Remote Sensing Change Detection

DGX agent

arXiv:2604.11231v1 Announce Type: new Abstract: Change detection is a fundamental task in remote sensing, aiming to quantify the impacts of human activities and ecological dynamics on land-cover chang

model-releasesarxiv-cs-cv
14 Apr 2026
Research

Self-Calibrating Language Models via Test-Time Discriminative Distillation

DGX agent

arXiv:2604.09624v1 Announce Type: new Abstract: Large language models (LLMs) are systematically overconfident: they routinely express high certainty on questions they often answer incorrectly. Existin

researcharxiv-cs-cl
14 Apr 2026
Research

SinkTrack: Attention Sink based Context Anchoring for Large Language Models

DGX agent

arXiv:2604.10027v1 Announce Type: new Abstract: Large language models (LLMs) suffer from hallucination and context forgetting. Prior studies suggest that attention drift is a primary cause of these pr

researcharxiv-cs-cv
14 Apr 2026
Local Ai

Stop burning GPU time on the wrong model. Manifest now supports Ollama Cloud 🦙🦚

DGX agent

Manifest is a tool or platform that has added support for Ollama Cloud, enabling users to route AI model inference to cloud-hosted Ollama instances rather than relying solely on local GPU resources. T

local-air-ollama
14 Apr 2026
Research

TokUR: Token-Level Uncertainty Estimation for Large Language Model Reasoning

DGX agent

arXiv:2505.11737v4 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have demonstrated impressive capabilities, their output quality remains inconsistent across various applica

researcharxiv-cs-ai
14 Apr 2026
Research

Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies

DGX agent

arXiv:2603.27960v2 Announce Type: replace-cross Abstract: Although Large Vision Language Models (LVLMs) have demonstrated impressive multimodal reasoning capabilities, their scalability and deployment

researcharxiv-cs-cl
14 Apr 2026
Applications

Understanding Generalization in Role-Playing Models via Information Theory

DGX agent

arXiv:2512.17270v2 Announce Type: replace-cross Abstract: Role-playing models (RPMs) are widely used in real-world applications but underperform when deployed in the wild. This degradation can be attr

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

Users accuse Anthropic of degrading Claude Opus 4.6's and Claude Code's performance; the startup's employees publicly deny it degrades models to manage capacity (Carl Franzen/VentureBeat)

DGX agent

Carl Franzen / VentureBeat: Users accuse Anthropic of degrading Claude Opus 4.6's and Claude Code's performance; the startup's employees publicly deny it degrades models to manage capacity — A growing

model-releasestechmeme
14 Apr 2026
Tutorials

Using Deep Learning Models Pretrained by Self-Supervised Learning for Protein Localization

DGX agent

arXiv:2604.10970v1 Announce Type: new Abstract: Background: Task-specific microscopy datasets are often small, making it difficult to train deep learning models that learn robust features. While self-

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models

DGX agent

arXiv:2601.06165v2 Announce Type: replace-cross Abstract: Current vision-language benchmarks predominantly feature well-structured questions with clear, explicit prompts. However, real user queries ar

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

When simulations look right but causal effects go wrong: Large language models as behavioral simulators

DGX agent

arXiv:2604.02458v2 Announce Type: replace-cross Abstract: Behavioral simulation is increasingly used to anticipate responses to interventions. Large language models (LLMs) enable researchers to specif

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

YUV20K: A Complexity-Driven Benchmark and Trajectory-Aware Alignment Model for Video Camouflaged Object Detection

DGX agent

arXiv:2604.09985v1 Announce Type: new Abstract: Video Camouflaged Object Detection (VCOD) is currently constrained by the scarcity of challenging benchmarks and the limited robustness of models agains

model-releasesarxiv-cs-cv
14 Apr 2026
Research

A Predictive View on Streaming Hidden Markov Models

DGX agent

arXiv:2604.09208v1 Announce Type: cross Abstract: We develop a predictive-first optimisation framework for streaming hidden Markov models. Unlike classical approaches that prioritise full posterior re

researcharxiv-cs-lg
13 Apr 2026
Safety

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention

DGX agent

arXiv:2511.18960v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown remarkable progress in embodied tasks recently, but most methods process visual observations in

safetyarxiv-cs-cv
13 Apr 2026
Tutorials

CT-1: Vision-Language-Camera Models Transfer Spatial Reasoning Knowledge to Camera-Controllable Video Generation

DGX agent

arXiv:2604.09201v1 Announce Type: new Abstract: Camera-controllable video generation aims to synthesize videos with flexible and physically plausible camera movements. However, existing methods either

tutorialsarxiv-cs-cv
13 Apr 2026
Model Releases

Cybersecurity analysis: Claude Mythos Preview had a 73% success rate on expert-level capture-the-flag challenges, which no model could finish before April 2025 (AI Security Institute)

DGX agent

AI Security Institute: Cybersecurity analysis: Claude Mythos Preview had a 73% success rate on expert-level capture-the-flag challenges, which no model could finish before April 2025 — The AI Security

model-releasestechmeme
13 Apr 2026
Local Ai

Does anyone know which model and potentially Lora was used to create these?

DGX agent

This Reddit thread from r/StableDiffusion is a community-driven reverse-identification request, where a user shares AI-generated images and asks fellow community members to help determine which Stable

local-air-stablediffusion
13 Apr 2026
Model Releases

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks

DGX agent

arXiv:2604.09535v1 Announce Type: new Abstract: Large foundation models have made significant advances in embodied intelligence, enabling synthesis and reasoning over egocentric input for household ta

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models

DGX agent

arXiv:2503.14075v3 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) have demonstrated remarkable capabilities in open-world multimodal understanding, yet their high computati

researcharxiv-cs-cl
13 Apr 2026
Local Ai

I made a playable ping pong game where every frame is ai generated. This is my interactive diffusion model I made from scratch.

DGX agent

A Reddit user on r/StableDiffusion showcased a fully playable ping pong game in which every individual frame is rendered in real time by a custom-built interactive diffusion model, rather than using t

local-air-stablediffusion
13 Apr 2026
Model Releases

Inpaint workflows for z-image, qwen and flux fill onereward

DGX agent

This Reddit post from r/StableDiffusion shares ComfyUI inpainting workflows for several modern AI image models, including Z-Image, Qwen Image/Edit, and Flux-series models . Flux Fill is a dedicated in

model-releasesr-stablediffusion
13 Apr 2026
Model Releases

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

DGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

model-releasesarxiv-cs-ai
13 Apr 2026
Local Ai

llama4 108b

DGX agent

This Reddit thread on r/ollama discusses running Meta's Llama 4 Maverick — a ~108B parameter model — locally using Ollama. Llama 4 models are natively multimodal AI models supporting text and image un

local-air-ollama
13 Apr 2026
Tutorials

Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym

DGX agent

arXiv:2604.09338v1 Announce Type: new Abstract: Spatial reasoning is central to navigation and robotics, yet measuring model capabilities on these tasks remains difficult. Existing benchmarks evaluate

tutorialsarxiv-cs-ai
13 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Multi-Exposure Image Fusion in Dynamic Scenes (Track 2)

DGX agent

arXiv:2604.09030v1 Announce Type: new Abstract: This paper presents NTIRE 2026, the 3rd Restore Any Image Model (RAIM) challenge on multi-exposure image fusion in dynamic scenes. We introduce a benchm

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Quantisation Reshapes the Metacognitive Geometry of Language Models

DGX agent

arXiv:2604.08976v1 Announce Type: new Abstract: We report that model quantisation restructures domain-level metacognitive efficiency in LLMs rather than degrading it uniformly. Evaluating Llama-3-8B-I

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a…

DGX agent

Scaling model size is hitting diminishing returns. The real gains are in orchestration. Our Co-Founder & Co-CEO @yshoham makes the case in a rare long-form profile by @Calcalistech today. The man tryi

model-releasesai21-labs--x
13 Apr 2026
Safety

SCoRe: Clean Image Generation from Diffusion Models Trained on Noisy Images

DGX agent

arXiv:2604.09436v1 Announce Type: new Abstract: Diffusion models trained on noisy datasets often reproduce high-frequency training artifacts, significantly degrading generation quality. To address thi

safetyarxiv-cs-cv
13 Apr 2026
Model Releases

SessionIntentBench: A Multi-task Inter-session Intention-shift Modeling Benchmark for E-commerce Customer Behavior Understanding

DGX agent

arXiv:2507.20185v2 Announce Type: replace Abstract: Session history is a common way of recording user interacting behaviors throughout a browsing activity with multiple products. For example, if an us

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

DGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

safetyarxiv-cs-ai
13 Apr 2026
Research

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

DGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

researcharxiv-cs-ai
13 Apr 2026
Agents

Great breakdown of how model providers are platformizing their AI/agents. A lot of people will take the convenience of going all in on a pro…

DGX agent

Great breakdown of how model providers are platformizing their AI/agents. A lot of people will take the convenience of going all in on a provider, but they will be locked in and giving up data control

agentsharrison-chase--x
11 Apr 2026
← Previous
1…152153154155156…1262
Next →