AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
Model Releases

TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training

DGX agent

arXiv:2604.24088v1 Announce Type: cross Abstract: Handling communication overhead in large-scale tensor-parallel training remains a critical challenge due to the dense, near-zero distributions of inte

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Test of Time: Rethinking Temporal Signal of Benchmark Contamination

DGX agent

arXiv:2509.00072v3 Announce Type: replace Abstract: Post-cutoff performance decay has been widely interpreted as a temporal signal for benchmark contamination. We critically examine this belief and de

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

DGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

DGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers

DGX agent

arXiv:2604.24155v1 Announce Type: cross Abstract: The quest to align machine behavior with human values raises fundamental questions about the moral frameworks that should govern AI decision-making. M

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The California Coastal Commission has issued a formal apology to @elonmusk and SpaceX, adding that it will not consider political views or s…

DGX agent

The California Coastal Commission has issued a formal apology to @elonmusk and SpaceX, adding that it will not consider political views or speech in future regulatory decisions. • The Commission admit

model-releaseselon-musk--x
28 Apr 2026
Model Releases

The cracks inside OpenAI are deepening, and the numbers don’t lie. When your own CFO is sounding the alarm, something is seriously wrong. Ch…

DGX agent

The cracks inside OpenAI are deepening, and the numbers don’t lie. When your own CFO is sounding the alarm, something is seriously wrong. Check this out: 1: OpenAI missed its target of 1 billion weekl

model-releasesgary-marcus--x
28 Apr 2026
Model Releases

The FIDO Alliance launches two working groups to establish industry standards for securing AI agent transactions; Google contributes the Agent Payments Protocol (Lily Hay Newman/Wired)

DGX agent

Lily Hay Newman / Wired: The FIDO Alliance launches two working groups to establish industry standards for securing AI agent transactions; Google contributes the Agent Payments Protocol — AI agents ma

model-releasestechmeme
28 Apr 2026
Model Releases

The Kerimov-Alekberli Model: An Information-Geometric Framework for Real-Time System Stability

DGX agent

arXiv:2604.24083v1 Announce Type: new Abstract: This study introduces the Kerimov-Alekberli model, a novel information-geometric framework that redefines AI safety by formally linking non-equilibrium

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Optimal Sample Complexity of Multiclass and List Learning

DGX agent

arXiv:2604.24749v1 Announce Type: new Abstract: While the optimal sample complexity of binary classification in terms of the VC dimension is well-established, determining the optimal sample complexity

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation

DGX agent

arXiv:2604.23750v1 Announce Type: cross Abstract: Hypernetwork-based methods such as Doc-to-LoRA internalize a document into an LLM's weights in a single forward pass, but they fail systematically on

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference

DGX agent

arXiv:2604.24079v1 Announce Type: cross Abstract: Large Language Models (LLMs) reveal inherent and distinctive personas through dialogue. However, most existing persona discovery approaches rely on su

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

DGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Randomness Floor: Measuring Intrinsic Non-Randomness in Language Model Token Distributions

DGX agent

arXiv:2604.22771v1 Announce Type: cross Abstract: Language models cannot be random. This paper introduces Entropic Deviation (ED), the normalised KL divergence between a model's token distribution and

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Rise of Large Language Models and the Direction and Impact of US Federal Research Funding

DGX agent

arXiv:2601.15485v2 Announce Type: replace-cross Abstract: Federal research funding shapes the direction, diversity, and impact of the US scientific enterprise. Large language models (LLMs) are rapidly

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Security Cost of Intelligence: AI Capability, Cyber Risk, and Deployment Paradox

DGX agent

arXiv:2604.23058v1 Announce Type: cross Abstract: Firms are deploying more capable AI systems, but organizational controls often have not kept pace. These systems can generate greater productivity gai

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

DGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models

DGX agent

arXiv:2511.08577v2 Announce Type: replace-cross Abstract: Improving reasoning abilities of Large Language Models (LLMs), especially under parameter constraints, is crucial for real-world applications.

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

this feels like less a trend and just something we expect... but q1 had a lot of launches that overlapped with recent startups i've seen: - …

DGX agent

this feels like less a trend and just something we expect... but q1 had a lot of launches that overlapped with recent startups i've seen: - codex desktop for running parallel - managed agents by claud

model-releasesyohei-nakajima--x
28 Apr 2026
Model Releases

This is great - @deepseek_ai V4 supports prefill! :D Most other providers have been dropping support for this critically important capabilit…

DGX agent

This is great - @deepseek_ai V4 supports prefill! :D Most other providers have been dropping support for this critically important capability, so wonderful to see at least one company stepping up. htt

model-releasesjeremy-howard--x
28 Apr 2026
Model Releases

Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B active MoE model built for agentic coding and l…

DGX agent

Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B active MoE model built for agentic coding and long-horizon tasks. Trained fully in-house on our own stack.

model-releasesclem-delangue--x
28 Apr 2026
Model Releases

Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0

DGX agent

Together AI announced immediate availability of NVIDIA's Nemotron 3 Nano Omni model to developers through its platform on the day of its release. The Nemotron 3 Nano Omni is a lightweight multimodal m

model-releasestogether-ai-blog
28 Apr 2026
Model Releases

TOL: Textual Localization with OpenStreetMap

DGX agent

arXiv:2604.01644v2 Announce Type: replace Abstract: Natural language provides an intuitive way to express spatial intent in geospatial applications. While existing localization methods often rely on d

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

TopoHR: Hierarchical Centerline Representation for Cyclic Topology Reasoning in Driving Scenes with Point-to-Instance Relations

DGX agent

arXiv:2604.24119v1 Announce Type: new Abstract: Topology reasoning is crucial for autonomous driving. Current methods primarily focus on instance-level learning for centerline detection, followed by a

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Toward Polymorphic Backdoor against Semantic Communication via Intensity-Based Poisoning

DGX agent

arXiv:2604.23231v1 Announce Type: cross Abstract: Semantic Communication (SC) backdoor attacks aim to utilize triggers to manipulate the system into producing predetermined outputs via backdoored shar

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Towards Continual Expansion of Data Coverage: Automatic Text-guided Edge-case Synthesis

DGX agent

arXiv:2509.26158v2 Announce Type: replace-cross Abstract: The performance of deep neural networks is strongly influenced by the quality of their training data. However, mitigating dataset bias by manu

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

DGX agent

arXiv:2603.25158v4 Announce Type: replace Abstract: Equipping Large Language Model (LLM) agents with domain-specific skills is critical for tackling complex tasks. Yet, manual authoring creates a seve

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Try now: https://www.together.ai/models/nvidia-nemotron-3-nano-omni#

DGX agent

NVIDIA Nemotron-3 Nano Omni is now available to try through Together AI's platform, offering access to a compact multimodal model capable of processing both text and audio inputs. This announcement hi

model-releasestogether-ai--x
28 Apr 2026
Model Releases

Tube Diffusion Policy: Reactive Visual-Tactile Policy Learning for Contact-rich Manipulation

DGX agent

arXiv:2604.23609v1 Announce Type: new Abstract: Contact-rich manipulation is central to many everyday human activities, requiring continuous adaptation to contact uncertainty and external disturbances

model-releasesarxiv-cs-ro
28 Apr 2026
Model Releases

Ulterior Motives: Detecting Misaligned Reasoning in Continuous Thought Models

DGX agent

arXiv:2604.23460v1 Announce Type: new Abstract: Chain-of-Thought (CoT) reasoning has emerged as a key technique for eliciting complex reasoning in Large Language Models (LLMs). Although interpretable,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Understanding the Limits of Automated Evaluation for Code Review Bots in Practice

DGX agent

arXiv:2604.24525v1 Announce Type: cross Abstract: Automated code review (ACR) bots are increasingly used in industrial software development to assist developers during pull request (PR) review. As ado

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks

DGX agent

arXiv:2604.23145v1 Announce Type: cross Abstract: Video Question Answering (VideoQA) demands models that jointly reason over spatial, temporal, and linguistic cues. However, the task's inherent comple

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

US startup Poolside debuts its first open-weight model, Laguna XS.2, a 33B-A3B-parameter MoE model, and Laguna M.1, a proprietary 225B-A23B-parameter MoE model (Carl Franzen/VentureBeat)

DGX agent

Carl Franzen / VentureBeat: US startup Poolside debuts its first open-weight model, Laguna XS.2, a 33B-A3B-parameter MoE model, and Laguna M.1, a proprietary 225B-A23B-parameter MoE model — The AI rac

model-releasestechmeme
28 Apr 2026
Model Releases

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

DGX agent

arXiv:2509.14837v2 Announce Type: replace Abstract: Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mec

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

VAPO: End-to-end Slide-Enhanced Speech Recognition with Omni-modal Large Language Models

DGX agent

arXiv:2510.08618v2 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) offer a promising end-to-end solution for slide-enhanced speech recognition due to their inherent mul

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Want to see which frontier models do the best on document understanding? Check out our ParseBench leaderboard on @kaggle! https://www.kaggle…

DGX agent

Want to see which frontier models do the best on document understanding? Check out our ParseBench leaderboard on @kaggle! https://www.kaggle.com/benchmarks/llamaindex-org/parsebench For more details o

model-releasesjerry-liu--x
28 Apr 2026
Model Releases

WebSerial Vision Training for Microcontrollers: A Browser-Based Companion to On-Device CNN Training

DGX agent

arXiv:2604.22834v1 Announce Type: new Abstract: This paper presents webmcu-vision-web, a single-file, zero-install browser application for end-to-end TinyML vision model training and deployment on the

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Welcome to the agentic era: Public sector highlights and reflections from Next ‘26

DGX agent

Welcome to the agentic era! Last week, leaders from our public sector customer and partner ecosystem took the stage at Google Cloud Next to share how they are leveraging AI and agents to scale their i

model-releasesgoogle-cloud-ai
28 Apr 2026
Model Releases

we're doing a lot more of this, hunting down some of the most annoying bugs in Claude Code let me know if you have any white whales

DGX agent

we're doing a lot more of this, hunting down some of the most annoying bugs in Claude Code let me know if you have any white whales In the last four Claude Code CLI releases, we’ve shipped 50+ stabili

model-releasesthariq--x
28 Apr 2026
Model Releases

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

DGX agent

arXiv:2604.23942v1 Announce Type: cross Abstract: People increasingly turn to large language models (LLMs) to interpret ambiguous social situations: a delayed text reply, an unusually cold supervisor,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

What search providers are you all using with openclaw/pi/opencode/etc? Brave; serpapi; gemini; ...? Got any favorites?

DGX agent

This post is a question about search provider preferences for use with various open-source tools and APIs (openclaw, pi, opencode, etc.), with examples of options like Brave, SerpAPI, and Gemini. The

model-releasesjeremy-howard--x
28 Apr 2026
Model Releases

When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL

DGX agent

arXiv:2604.23398v1 Announce Type: new Abstract: We report a reproducible error pattern in GPT-5.4 on OWL~2~DL compliance queries: the model frequently answers ``unknown'' when the reasoner-entailed an

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

When Does Removing LayerNorm Help? Activation Bounding as a Regime-Dependent Implicit Regularizer

DGX agent

arXiv:2604.23434v1 Announce Type: cross Abstract: Dynamic Tanh (DyT) removes LayerNorm by bounding activations with a learned tanh(alpha x). We show that this bounding is a regime-dependent implicit r

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of Multi-line Handwritten Math OCR

DGX agent

arXiv:2604.22774v1 Announce Type: cross Abstract: Accurate transcription of handwritten mathematics is crucial for educational AI systems, yet current benchmarks fail to evaluate this capability prope

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

WISE-FM:Operation-Aware, Engineering-Informed Foundation Model for Multi-Task Well Design

DGX agent

arXiv:2604.23767v1 Announce Type: new Abstract: Deploying machine learning models across diverse well portfolios requires generalisation to wells with design parameters outside the training distributi

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Wordle 1,773 5/6 ⬛⬛⬛🟨🟩 ⬛🟨⬛⬛⬛ ⬛🟨🟨⬛🟩 ⬛⬛🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This is a Wordle game result posted by Anthropic on X (formerly Twitter), showing the solution was found in 5 out of 6 possible attempts. The emoji grid indicates the letter placement feedback for eac

model-releasesanthropic--x
28 Apr 2026
Model Releases

xOffense: An Autonomous Multi-Agent Framework for Penetration Testing with Domain-Adapted Large Language Models

DGX agent

arXiv:2509.13021v2 Announce Type: replace-cross Abstract: This work introduces xOffense, an AI-driven, multi-agent penetration testing framework that shifts the process from labor-intensive, expert-dr

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Yes this is a line from the system prompt.

DGX agent

Yes this is a line from the system prompt. gpt-5.5 prompt for codex seems to have a duplicated line trying to get it to not talk about creatures? Never talk about goblins, gremlins, raccoons, trolls,

model-releasesethan-mollick--x
28 Apr 2026
← Previous
1…387388389390391…470
Next →