DEV Community

Programming

The magic behind computers. 💻 🪄

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
AI Automates MOOSE Simulations: Introducing the MooseAgent Framework

AI Automates MOOSE Simulations: Introducing the MooseAgent Framework

Comments
8 min read
LLM Search Boost: ReZero Rewards Retry After Initial RAG Failures

LLM Search Boost: ReZero Rewards Retry After Initial RAG Failures

Comments
10 min read
SIFT-50M: New Data Supercharges Speech LLMs, Improves Understanding

SIFT-50M: New Data Supercharges Speech LLMs, Improves Understanding

Comments
6 min read
OctGPT: Autoregressive 3D Shape Generation Rivals Diffusion Models

OctGPT: Autoregressive 3D Shape Generation Rivals Diffusion Models

Comments
8 min read
VisuoThink: Visual-Text AI Beats Reasoning Limits with Multimodal Tree Search

VisuoThink: Visual-Text AI Beats Reasoning Limits with Multimodal Tree Search

Comments
13 min read
AI Team Cracks LLM Reasoning: New Model + "CEO" Agent Beats Benchmarks

AI Team Cracks LLM Reasoning: New Model + "CEO" Agent Beats Benchmarks

Comments
7 min read
SocioVerse: LLM Agents Simulate 10M Real Users for Realistic Social Behavior

SocioVerse: LLM Agents Simulate 10M Real Users for Realistic Social Behavior

Comments
8 min read
Mavors: Multi-Granularity Video Beats MLLM Limits! See How.

Mavors: Multi-Granularity Video Beats MLLM Limits! See How.

Comments
11 min read
Self-Training Boosts Code Generation: RewardRanker Outperforms GPT-3.5 & Rivals GPT-4

Self-Training Boosts Code Generation: RewardRanker Outperforms GPT-3.5 & Rivals GPT-4

Comments
5 min read
LLMs Can "Bleed" Knowledge: How New Data Warps AI Understanding

LLMs Can "Bleed" Knowledge: How New Data Warps AI Understanding

Comments
9 min read
VL-Rethinker: RL Drives Self-Reflection in Vision-Language Models for Smarter Reasoning

VL-Rethinker: RL Drives Self-Reflection in Vision-Language Models for Smarter Reasoning

Comments
9 min read
TextArena: LLM Games Test Reasoning, Negotiation, & Deception Skills

TextArena: LLM Games Test Reasoning, Negotiation, & Deception Skills

Comments
6 min read
MedHal: New Dataset Flags Medical AI Lies - Can AI Detect False Health Info?

MedHal: New Dataset Flags Medical AI Lies - Can AI Detect False Health Info?

Comments
7 min read
FUSION: Deep Vision-Language Integration Outperforms with Fewer Tokens

FUSION: Deep Vision-Language Integration Outperforms with Fewer Tokens

Comments
10 min read
Single Transformer Beats Modular Vision-Language Models in New Study

Single Transformer Beats Modular Vision-Language Models in New Study

Comments
10 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.