Apertium has moved from SourceForge to GitHub.
If you have any questions, please come and talk to us on #apertium on irc.freenode.net or contact the GitHub migration team.

User:Eden/GSoC progress

From Apertium
< User:Eden(Difference between revisions)
Jump to: navigation, search
(Status table)
Line 98: Line 98:
 
| 7
 
| 7
 
| July 8 - July 14
 
| July 8 - July 14
  +
|1,229
  +
|1,573
 
|
 
|
 
|
 
|
|
+
|68.11%,54.90%
|
+
|74.87%,61.73%
|
 
|
 
 
|
 
|
 
|
 
|

Revision as of 05:39, 11 July 2019

Status table

Week Stems naïve coverage WER,PER Progress
dates lin lin-eng lin lin-eng lin→eng eng→lin Evaluation Notes
0 May 20 - May 26 727 139 61.95% 40.86% 86.79%,80.87% 75.27%,63.98%
1 May 27 - June 02 904 139 62.57% 40.86% 86.79%,80.87% 75.27%,63.98%
2 May 03 - June 09 1,154 1,416 63.17% 53.03% 87.02%,79.95% 74.46%,60.22%
3 June 10 - June 16 1,172 1,501 61.60% 91.57%,79.04% 75.85%,62.90% WER for 'lin-eng' went up because of an incomplete rule for verbs that creates unnecessary pronouns. Main work next week will be on rules to dramatically improve WER and PER.
4 June 17 - June 23 1,200 1,540 69.70% 62.70% 79.27%,64.24% 84.41%,72.58%
5 June 24 - June 30 1,200 1,556 70.21% 61.90% 77.68%,67.88% 85.48%,73.92%
6 July 1 - July 7
7 July 8 - July 14 1,229 1,573 68.11%,54.90% 74.87%,61.73%

Notes

  • To count stems in lexc, try:
 grep -E ":\w+.*;" apertium-lin.lin.lexc | grep -v "[<>]" | wc -l
  • To count stems in the bidix, try this:
 grep "<p" apertium-eng-lin.eng-lin.dix  | wc -l
  • To get WER and PER use apertium-eval-translator-line
Personal tools