I shudder at the complexity of the task, but I almost think you need to manually tweak the vector space of a model that already works, and I have no idea how to do that in practice.
The amount of text required for a machine to grind through it millions of times to tease out the shape of a language doesn't sound like something you have. If you have the time of native speakers, it might be possible to build tools for them to correct the most "off" parts of the model interactively.
Comments
I shudder at the complexity of the task, but I almost think you need to manually tweak the vector space of a model that already works, and I have no idea how to do that in practice.
The amount of text required for a machine to grind through it millions of times to tease out the shape of a language doesn't sound like something you have. If you have the time of native speakers, it might be possible to build tools for them to correct the most "off" parts of the model interactively.