AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Local Ai

Unrealized Expectations: Comparing AI Methods vs Classical Algorithms for Maximum Independent Set

DGX agent

arXiv:2502.03669v3 Announce Type: replace-cross Abstract: AI methods, such as generative models and reinforcement learning, have recently been applied to combinatorial optimization (CO) problems, espe

local-aiarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

DGX agent

arXiv:2604.23674v1 Announce Type: new Abstract: With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people co

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

WebSerial Vision Training for Microcontrollers: A Browser-Based Companion to On-Device CNN Training

DGX agent

arXiv:2604.22834v1 Announce Type: new Abstract: This paper presents webmcu-vision-web, a single-file, zero-install browser application for end-to-end TinyML vision model training and deployment on the

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

DGX agent

arXiv:2604.23942v1 Announce Type: cross Abstract: People increasingly turn to large language models (LLMs) to interpret ambiguous social situations: a delayed text reply, an unusually cold supervisor,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL

DGX agent

arXiv:2604.23398v1 Announce Type: new Abstract: We report a reproducible error pattern in GPT-5.4 on OWL~2~DL compliance queries: the model frequently answers ``unknown'' when the reasoner-entailed an

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Adapting MLLMs for Nuanced Video Retrieval

DGX agent

arXiv:2512.13511v2 Announce Type: replace Abstract: Our objective is to build an embedding model that captures the nuanced relationship between a search query and candidate videos. We cover three aspe

researcharxiv-cs-cv
27 Apr 2026
Safety

AgentBound: Securing Execution Boundaries of AI Agents

DGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

safetyarxiv-cs-ai
27 Apr 2026
Model Releases

Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography

DGX agent

arXiv:2604.22139v1 Announce Type: new Abstract: Reliable automated analysis of Optical Coherence Tomography (OCT) imaging is crucial for diagnosing retinal disorders but faces a critical barrier: the

model-releasesarxiv-cs-cv
27 Apr 2026
Research

Chain-of-Memory: Lightweight Memory Construction with Dynamic Evolution for LLM Agents

DGX agent

arXiv:2601.14287v2 Announce Type: replace Abstract: External memory systems are pivotal for enabling Large Language Model (LLM) agents to maintain persistent knowledge and perform long-horizon decisio

researcharxiv-cs-lg
27 Apr 2026
Tutorials

CLVAE: A Variational Autoencoder for Long-Term Customer Revenue Forecasting

DGX agent

arXiv:2604.22636v1 Announce Type: cross Abstract: Predicting customers' long-term revenue from sparse and irregular transaction data is central to marketing resource allocation in non-contractual sett

tutorialsarxiv-cs-lg
27 Apr 2026
Safety

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

DGX agent

arXiv:2604.22542v1 Announce Type: cross Abstract: Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. T

safetyarxiv-cs-ai
27 Apr 2026
Research

Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation

DGX agent

arXiv:2510.19592v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) demonstrate strong video understanding by attending to visual tokens relevant to textual queries. To direct

researcharxiv-cs-cv
27 Apr 2026
Model Releases

EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges

DGX agent

arXiv:2604.22595v1 Announce Type: new Abstract: CLIP has demonstrated strong generalization in visual domains through natural language supervision, even for video action recognition. However, most exi

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

@FireworksAI_HQ deeply believes in delivering the frontier quality. We will spend all the effort giving our users and the broad community th…

DGX agent

@FireworksAI_HQ deeply believes in delivering the frontier quality. We will spend all the effort giving our users and the broad community the best OSS model quality. Deepseek V4 API up anytime now aft

model-releasesfireworks-ai--x
27 Apr 2026
Agents

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we exp…

DGX agent

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we explored what happens if we train an AI to do that job instead.

agentsdavid-ha--x
27 Apr 2026
Safety

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

DGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

safetyarxiv-cs-cv
27 Apr 2026
Model Releases

how to adjust the thinking effort for deepseek v4 on ollama cloud

DGX agent

DeepSeek V4 models on Ollama Cloud support three thinking modes: 'No thinking' for fast answers, 'Thinking' for careful analysis, and 'Max thinking' for maximum reasoning effort . Users can adjust thi

model-releasesr-ollama
27 Apr 2026
Research

LayerBoost: Layer-Aware Attention Reduction for Efficient LLMs

DGX agent

arXiv:2604.22050v1 Announce Type: cross Abstract: Transformers are mostly relying on softmax attention, which introduces quadratic complexity with respect to sequence length and remains a major bottle

researcharxiv-cs-cl
27 Apr 2026
Model Releases

LLMs as Assessors: Right for the Right Reason?

DGX agent

arXiv:2601.08919v2 Announce Type: replace-cross Abstract: A good deal of recent research has focused on how Large Language Models (LLMs) may be used as judges in place of humans to evaluate the qualit

model-releasesarxiv-cs-cl
27 Apr 2026
Research

NiuTrans.LMT: Toward Inclusive and Scalable Multilingual Machine Translation with LLMs

DGX agent

arXiv:2511.07003v2 Announce Type: replace Abstract: Large language models have significantly advanced Multilingual Machine Translation (MMT), yet scaling to many languages while keeping quality robust

researcharxiv-cs-cl
27 Apr 2026
Model Releases

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

DGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

PL-MTEB: Polish Massive Text Embedding Benchmark

DGX agent

arXiv:2405.10138v2 Announce Type: replace Abstract: In this paper, we introduce the Polish Massive Text Embedding Benchmark (PL-MTEB), a comprehensive benchmark for text embeddings in the Polish langu

model-releasesarxiv-cs-cl
27 Apr 2026
Research

PreMoE: Proactive Inference for Efficient Mixture-of-Experts

DGX agent

arXiv:2505.17639v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models offer dynamic computation, but are typically deployed as static full-capacity models, missing opportunities for depl

researcharxiv-cs-lg
27 Apr 2026
Research

PrivUn: Unveiling Latent Ripple Effects and Shallow Forgetting in Privacy Unlearning

DGX agent

arXiv:2604.22076v1 Announce Type: cross Abstract: Large language models (LLMs) often memorize private information during training, raising serious privacy concerns. While machine unlearning has emerge

researcharxiv-cs-cl
27 Apr 2026
Model Releases

Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition

DGX agent

arXiv:2604.22390v1 Announce Type: new Abstract: Visual Place Recognition (VPR) determines a query image's geographic location by matching it against geotagged databases. However, existing methods stru

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

DGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Rethinking Math Reasoning Evaluation: A Robust LLM-as-a-Judge Framework Beyond Symbolic Rigidity

DGX agent

arXiv:2604.22597v1 Announce Type: new Abstract: Recent advancements in large language models have led to significant improvements across various tasks, including mathematical reasoning, which is used

researcharxiv-cs-ai
27 Apr 2026
Applications

Segment Any-Quality Images with Generative Latent Space Enhancement

DGX agent

arXiv:2503.12507v3 Announce Type: replace Abstract: Despite their success, Segment Anything Models (SAMs) experience significant performance drops on severely degraded, low-quality images, limiting th

applicationsarxiv-cs-cv
27 Apr 2026
Research

SSG: Logit-Balanced Vocabulary Partitioning for LLM Watermarking

DGX agent

arXiv:2604.22438v1 Announce Type: cross Abstract: Watermarking has emerged as a promising technique for tracing the authorship of content generated by large language models (LLMs). Among existing appr

researcharxiv-cs-ai
27 Apr 2026
Industry

This is how it's done! Who else should we ask to release weights?

DGX agent

Clem Delangue, CEO of Hugging Face, discusses best practices for releasing model weights in the open-source AI community and advocates for other AI organizations to follow suit in making their models

industryclem-delangue--x
27 Apr 2026
Safety

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem

DGX agent

arXiv:2506.17299v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become increasingly deployed in safety-critical applications, the lack of systematic methods to assess their v

safetyarxiv-cs-ai
27 Apr 2026
Research

Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations

DGX agent

arXiv:2601.03779v2 Announce Type: replace Abstract: We explore intrinsic dimension (ID) of LLM representations as a marker of linguistic complexity. Specifically, we test whether ID differences across

researcharxiv-cs-cl
27 Apr 2026
Model Releases

UR^2: Unify RAG and Reasoning through Reinforcement Learning

DGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threa…

DGX agent

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threads. I don’t think we’ve cracked the right UI for managing ag

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate'…

DGX agent

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image

model-releasesallie-k--miller--x
26 Apr 2026
Model Releases

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST t…

DGX agent

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST time I feel I have a frontier model running on my computer. T

model-releasesclem-delangue--x
26 Apr 2026
Industry

5.5 is so earnest 'little engine that could' energy

DGX agent

Sam Altman praised OpenAI's o1 model (version 5.5) for its earnest, persistent approach to problem-solving, comparing it to the 'little engine that could' mentality. The comment reflects Altman's pers

industrysam-altman--x
25 Apr 2026
Model Releases

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹C…

DGX agent

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹Claude Code: Set model to deepseek-v4-pro[1m] to unlock 1M conte

model-releasesdeepseek--x
25 Apr 2026
Research

http://reddit.com/r/LocalLLaMA

DGX agent

r/LocalLLaMA is a subreddit community dedicated to discussing and sharing resources about running large language models locally on personal computers, covering topics like model optimization, hardware

researchnous-research--x
25 Apr 2026
Model Releases

WHY ARE YOU LIKE THIS

DGX agent

@scottjla on Twitter in reply to my pelican riding a bicycle benchmark: I feel like we need to stack these tests now I checked to confirm that the model (ChatGPT Images 2.0) added the 'WHY ARE YOU LIK

model-releasessimon-willison
25 Apr 2026
Model Releases

4⃣4⃣4⃣4⃣

DGX agent

4⃣4⃣4⃣4⃣ Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit

model-releasestogether-ai--x
24 Apr 2026
Model Releases

A-THENA: Early Intrusion Detection for IoT with Time-Aware Hybrid Encoding and Network-Specific Augmentation

DGX agent

arXiv:2604.21623v1 Announce Type: cross Abstract: The proliferation of Internet of Things (IoT) devices has significantly expanded attack surfaces, making IoT ecosystems particularly susceptible to so

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Absorber LLM: Harnessing Causal Synchronization for Test-Time Training

DGX agent

arXiv:2604.20915v1 Announce Type: cross Abstract: Transformers suffer from a high computational cost that grows with sequence length for self-attention, making inference in long streams prohibited by

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security

DGX agent

arXiv:2601.18491v2 Announce Type: replace Abstract: The rise of AI agents introduces complex safety and security challenges arising from autonomous tool use and environmental interactions. Current gua

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents

DGX agent

arXiv:2604.21134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently misread values, hallucinate details, and confuse overlapping elements in charts. Current approaches rely solely

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

BioMiner: A Multi-modal System for Automated Mining of Protein-Ligand Bioactivity Data from Literature

DGX agent

arXiv:2604.21508v1 Announce Type: new Abstract: Protein-ligand bioactivity data published in the literature are essential for drug discovery, yet manual curation struggles to keep pace with rapidly gr

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion

DGX agent

arXiv:2604.04395v2 Announce Type: replace Abstract: 3D conducting motion generation aims to synthesize fine-grained conductor motions from music, with broad potential in music education, virtual perfo

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

DGX agent

arXiv:2604.21308v1 Announce Type: cross Abstract: Enterprise LLM agents can dramatically improve workplace productivity, but their core capability, retrieving and using internal context to act on a us

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…560561562563564…1395
Next →