Often the problem with .pdf that parts of it (important text) are not grouped or structured. It would be nice to see this model being applied to such files to reconstruct a logically connected, grouped and structured text. I think that this would be a big deal.
Often the problem with .pdf that parts of it (important text) are not grouped or structured. It would be nice to see this model being applied to such files to reconstruct a logically connected, grouped and structured text. I think that this would be a big deal.