#machine-learning

Machine Learning

129 stories tagged Machine Learning

S2T-RLHF: Hierarchical Credit Assignment Improves Stability of Preference-Based RLHF Training
research

S2T-RLHF: Hierarchical Credit Assignment Improves Stability of Preference-Based RLHF Training

A new arXiv paper introduces S2T-RLHF, a method that uses hierarchical credit assignment to stabilize reinforcement learning from human feedback (RLHF). By breaking sequence-level rewards into finer token-level supervision, the approach reduces training instability and helps AI models learn human preferences more accurately, leading to more reliable AI assistants and tools.

Photoroom Releases PRX Part 4: An Open-Source Data Strategy for AI Training
open-source

Photoroom Releases PRX Part 4: An Open-Source Data Strategy for AI Training

Photoroom, the AI-powered photo editing startup, has published PRX Part 4, a detailed blog post on Hugging Face outlining its open-source data strategy for training high-quality AI models. The post reveals how the company curates, filters, and augments training data to achieve state-of-the-art results while keeping the approach transparent and reproducible.

via Hugging Face Blog#ai#open-source#data-strategy