参考:
de-fr
使用以下命令在 TFDS 中加载此数据集:
ds = tfds.load('huggingface:ecb/de-fr')
- 说明:
Original source: Website and documentatuion from the European Central Bank, compiled and made available by Alberto Simoes (thank you very much!)
19 languages, 170 bitexts
total number of files: 340
total number of tokens: 757.37M
total number of sentence fragments: 30.55M
- 许可:无已知许可
- 版本:1.0.0
- 拆分:
拆分 | 样本 |
---|---|
'train' |
105116 |
- 特征:
{
"id": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"translation": {
"languages": [
"de",
"fr"
],
"id": null,
"_type": "Translation"
}
}
cs-en
使用以下命令在 TFDS 中加载此数据集:
ds = tfds.load('huggingface:ecb/cs-en')
- 说明:
Original source: Website and documentatuion from the European Central Bank, compiled and made available by Alberto Simoes (thank you very much!)
19 languages, 170 bitexts
total number of files: 340
total number of tokens: 757.37M
total number of sentence fragments: 30.55M
- 许可:无已知许可
- 版本:1.0.0
- 拆分:
拆分 | 样本 |
---|---|
'train' |
63716 |
- 特征:
{
"id": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"translation": {
"languages": [
"cs",
"en"
],
"id": null,
"_type": "Translation"
}
}
el-it
使用以下命令在 TFDS 中加载此数据集:
ds = tfds.load('huggingface:ecb/el-it')
- 说明:
Original source: Website and documentatuion from the European Central Bank, compiled and made available by Alberto Simoes (thank you very much!)
19 languages, 170 bitexts
total number of files: 340
total number of tokens: 757.37M
total number of sentence fragments: 30.55M
- 许可:无已知许可
- 版本:1.0.0
- 拆分:
拆分 | 样本 |
---|---|
'train' |
94712 |
- 特征:
{
"id": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"translation": {
"languages": [
"el",
"it"
],
"id": null,
"_type": "Translation"
}
}
en-nl
使用以下命令在 TFDS 中加载此数据集:
ds = tfds.load('huggingface:ecb/en-nl')
- 说明:
Original source: Website and documentatuion from the European Central Bank, compiled and made available by Alberto Simoes (thank you very much!)
19 languages, 170 bitexts
total number of files: 340
total number of tokens: 757.37M
total number of sentence fragments: 30.55M
- 许可:无已知许可
- 版本:1.0.0
- 拆分:
拆分 | 样本 |
---|---|
'train' |
126482 |
- 特征:
{
"id": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"translation": {
"languages": [
"en",
"nl"
],
"id": null,
"_type": "Translation"
}
}
fi-pl
使用以下命令在 TFDS 中加载此数据集:
ds = tfds.load('huggingface:ecb/fi-pl')
- 说明:
Original source: Website and documentatuion from the European Central Bank, compiled and made available by Alberto Simoes (thank you very much!)
19 languages, 170 bitexts
total number of files: 340
total number of tokens: 757.37M
total number of sentence fragments: 30.55M
- 许可:无已知许可
- 版本:1.0.0
- 拆分:
拆分 | 样本 |
---|---|
'train' |
41686 |
- 特征:
{
"id": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"translation": {
"languages": [
"fi",
"pl"
],
"id": null,
"_type": "Translation"
}
}