DEV Community

Beginners

"A journey of a thousand miles begins with a single step." -Chinese Proverb

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
ARAIDA: Analogical Reasoning-Augmented Interactive Data Annotation

ARAIDA: Analogical Reasoning-Augmented Interactive Data Annotation

Comments
4 min read
Attention as an RNN

Attention as an RNN

Comments
4 min read
Representation noising effectively prevents harmful fine-tuning on LLMs

Representation noising effectively prevents harmful fine-tuning on LLMs

Comments
5 min read
ColorFoil: Investigating Color Blindness in Large Vision and Language Models

ColorFoil: Investigating Color Blindness in Large Vision and Language Models

Comments
4 min read
Evaluation of the Programming Skills of Large Language Models

Evaluation of the Programming Skills of Large Language Models

Comments
3 min read
The CAP Principle for LLM Serving

The CAP Principle for LLM Serving

Comments
4 min read
From Sparse to Soft Mixtures of Experts

From Sparse to Soft Mixtures of Experts

Comments
4 min read
Pareto Optimal Learning for Estimating Large Language Model Errors

Pareto Optimal Learning for Estimating Large Language Model Errors

Comments
4 min read
Transformers Can Do Arithmetic with the Right Embeddings

Transformers Can Do Arithmetic with the Right Embeddings

Comments
4 min read
Can a Transformer Represent a Kalman Filter?

Can a Transformer Represent a Kalman Filter?

Comments
4 min read
Training Language Models to Generate Text with Citations via Fine-grained Rewards

Training Language Models to Generate Text with Citations via Fine-grained Rewards

Comments
3 min read
VecFusion: Vector Font Generation with Diffusion

VecFusion: Vector Font Generation with Diffusion

1
Comments
4 min read
Thermodynamic Natural Gradient Descent

Thermodynamic Natural Gradient Descent

Comments
5 min read
Levels of AGI for Operationalizing Progress on the Path to AGI

Levels of AGI for Operationalizing Progress on the Path to AGI

Comments
4 min read
Why are Sensitive Functions Hard for Transformers?

Why are Sensitive Functions Hard for Transformers?

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.