Skip to content

Support power of 2 scaling factors in float8 training #2526

Support power of 2 scaling factors in float8 training

Support power of 2 scaling factors in float8 training #2526