Open X-Embodiment
1M+ real-robot trajectories from 20+ institutions.
Dataset Description A small subset of the-stack dataset, with 87 programming languages, each has 10,000 random samples from the original dataset.…
Dataset Description A small subset of the-stack dataset, with 87 programming languages, each has 10,000 random samples from the original dataset. Languages The dataset contains 87 programming languages: ‘ada’, ‘agda’, ‘alloy’, ‘antlr’, ‘applescript’, ‘assembly’, ‘augeas’, ‘awk’, ‘batchfile’, ‘bison’, ‘bluespec’, ‘c’, ‘c++’, ‘c-sharp’, ‘clojure’, ‘cmake’, ‘coffeescript’, ‘common-lisp’, ‘css’, ‘cuda’, ‘dart’, ‘dockerfile’, ‘elixir’, ‘elm’, ’emacs-lisp’,’erlang’… See the full description on the dataset page:
Source: Hugging Face Hub (bigcode/the-stack-smol-xl). Metadata imported from the dataset’s Hub tags.