You can not select more than 25 topics Topics must start with a letter or number, can include dashes ('-') and can be up to 35 characters long.
LSmyrnaios ff46839158 Fix not prioritizing the gradle version defined inside the "" script. 3 days ago
gradle/wrapper - Workaround a bug of Impala-JDBC-Driver, when creating insert-prepared-statements. 1 month ago
scripts Initial commit of UrlsController. 10 months ago
src/main - Make sure the temp table "current_assignment" from a cancelled previous execution, is dropped and purged on startup. 5 days ago - Implement the "getAndUploadFullTexts" functionality. In order to access the S3-ObjectStore from one trusted place, the Controller will request the files from the workers and upload them on S3. Afterwards, the workers will delete those files from their local storage. Previously, each worker uploaded its own files. 2 months ago
build.gradle Update dependencies. 3 days ago Fix not prioritizing the gradle version defined inside the "" script. 3 days ago
settings.gradle - Add the "isControllerAlive"-endpoint. 4 months ago


This is the Controller's Application.
It receives requests coming from the workers , constructs an assignments-list with data received from a database and returns the list to the workers.
Then it receives the "WorkerReports" and writes them into the database.
The database used is the Impala .

To install and run the application, run git clone. Then, provide a file "S3_minIO_credentials.txt", inside the working directory.
In the "S3_minIO_credentials.txt" file, you should provide the endpoint, the accessKey, the secretKey, the region and the bucket, in that order, separated by comma.
Afterwards, execute the script.
If you want to just run the app, then run the script with the argument "1": ./ 1.