Oliver Sieberling's picture

2 9 1

Oliver Sieberling

OliverSieberling

·

AI & ML interests

None yet

Recent Activity

upvoted a paper 11 days ago

Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning

upvoted a paper about 1 month ago

WUSH: Near-Optimal Adaptive Transforms for LLM Quantization

upvoted a paper 5 months ago

The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm

View all activity

Organizations

upvoted a paper 11 days ago

Emergent temporal abstractions in autoregressive models enable hierarchical reinforcement learning

Paper • 2512.20605 • Published 13 days ago • 60

upvoted a paper about 1 month ago

WUSH: Near-Optimal Adaptive Transforms for LLM Quantization

Paper • 2512.00956 • Published Nov 30, 2025 • 20

upvoted a paper 5 months ago

The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm

Paper • 2507.18553 • Published Jul 24, 2025 • 40

upvoted 3 papers 7 months ago

Unified Scaling Laws for Compressed Representations

Paper • 2506.01863 • Published Jun 2, 2025 • 19

SVD-Free Low-Rank Adaptive Gradient Optimization for Large Language Models

Paper • 2505.17967 • Published May 23, 2025 • 17

Quartet: Native FP4 Training Can Be Optimal for Large Language Models

Paper • 2505.14669 • Published May 20, 2025 • 78

authored a paper 7 months ago

Quartet: Native FP4 Training Can Be Optimal for Large Language Models

Paper • 2505.14669 • Published May 20, 2025 • 78

upvoted 2 papers 11 months ago

DarwinLM: Evolutionary Structured Pruning of Large Language Models

Paper • 2502.07780 • Published Feb 11, 2025 • 18

QuEST: Stable Training of LLMs with 1-Bit Weights and Activations

Paper • 2502.05003 • Published Feb 7, 2025 • 42

commented a paper about 1 year ago

Cottention: Linear Transformers With Cosine Attention

Paper • 2409.18747 • Published Sep 27, 2024 • 16 •

authored 3 papers about 1 year ago

EvoPress: Towards Optimal Dynamic Model Compression via Evolutionary Search

Paper • 2410.14649 • Published Oct 18, 2024 • 8

Plus Strategies are Exponentially Slower for Planted Optima of Random Height

Paper • 2404.09687 • Published Apr 15, 2024

Hardest Monotone Functions for Evolutionary Algorithms

Paper • 2311.07438 • Published Nov 13, 2023

liked a dataset over 1 year ago

cerebras/SlimPajama-627B

Preview • Updated Jul 7, 2023 • 57.6k • 511