On robot deployment flywheels and why I am skeptical of the data they produce
Great article! Given the need for novelty, what are your thoughts on actively training policies designed to target noisy/difficult distributions as supplementary data generators for the main policy (ex: seeking variance-seeking behavior)?
Interesting article...
Great article! Given the need for novelty, what are your thoughts on actively training policies designed to target noisy/difficult distributions as supplementary data generators for the main policy (ex: seeking variance-seeking behavior)?
Interesting article...