Commit Graph

8 Commits

Author SHA1 Message Date
97305d27ff Integrate support for more files into processing and upload
The restriction that only pdf files can be uploaded is removed. All
files can now be uploaded. The processing may not process all. It is
still possible to restrict file uploads by types via a configuration.
2020-02-19 23:27:00 +01:00
9b1349734e Convert some files to pdf 2020-02-19 02:03:10 +01:00
5869e2ee6e Streamline extern-conv stdin/infile 2020-02-18 12:43:47 +01:00
0dcc00836b Make logger configurable in system commands 2020-02-18 12:02:43 +01:00
e0682464b5 Configure pdf extraction; move Logger and DataType to common 2020-02-17 14:01:36 +01:00
3d615181e0 Early draft for text extraction 2020-02-17 01:57:22 +01:00
8143a4edcc Adding extraction primitives 2020-02-16 21:37:26 +01:00
851ee7ef0f Reorganize processing code
Use separate modules for

- text extraction
- conversion to pdf
- text analysis
2020-02-15 21:25:25 +01:00