Miriam Baglioni
|
588aca5ce4
|
Merge pull request 'h2020classification' (#280) from h2020classification into beta
Reviewed-on: #280
|
2023-03-03 09:29:10 +01:00 |
Claudio Atzori
|
8ec0d62d91
|
pre-group the records in each table before joning the contents from BETA and PROD together
|
2023-03-02 14:49:19 +01:00 |
Miriam Baglioni
|
0fff98a14c
|
[ECclassification] removed print
|
2023-03-02 11:46:57 +01:00 |
Miriam Baglioni
|
b0c2f7e526
|
[ECclassification] removed not needed resources
|
2023-03-02 11:44:48 +01:00 |
Miriam Baglioni
|
d4fc62c2f6
|
mergin with branch beta
|
2023-03-02 11:14:54 +01:00 |
Miriam Baglioni
|
de8ad1caef
|
[ECclassification] new implementation for the H2020 classification
|
2023-03-02 11:14:03 +01:00 |
Claudio Atzori
|
db9dad4aa7
|
[actionmanager] increased spark.sql.shuffle.partitions for publication, dataset, relation records
|
2023-03-02 09:11:37 +01:00 |
Miriam Baglioni
|
c1f9848953
|
[ECclassification] added new classes
|
2023-03-01 15:29:11 +01:00 |
Claudio Atzori
|
6f488547a7
|
ignore non processable records
|
2023-03-01 14:49:51 +01:00 |
Claudio Atzori
|
7d263f265e
|
adjusted logs
|
2023-03-01 11:58:07 +01:00 |
Claudio Atzori
|
16ad42e8f3
|
code formatting
|
2023-03-01 10:22:13 +01:00 |
Claudio Atzori
|
9c59dac859
|
followup changes reorganising the mdstore synchronisation mechanism
|
2023-03-01 10:16:20 +01:00 |
Miriam Baglioni
|
ad745c0aa3
|
[CrossrefFunderMapping] fixed issueson funder name
|
2023-02-28 14:58:27 +01:00 |
Miriam Baglioni
|
4f2df876cd
|
[ECclassification] new implementation first try
|
2023-02-28 14:44:00 +01:00 |
Claudio Atzori
|
2f7346e9cf
|
WIP monodirectional citations, Datacite
|
2023-02-28 13:30:51 +01:00 |
Claudio Atzori
|
0559d8b412
|
WIP monodirectional citations
|
2023-02-28 10:57:32 +01:00 |
Sandro La Bruzzo
|
69fa616490
|
removed wrong content
|
2023-02-28 10:27:38 +01:00 |
Sandro La Bruzzo
|
832a75d012
|
added mapping for crossref funder
|
2023-02-28 10:16:34 +01:00 |
Sandro La Bruzzo
|
78e51c182a
|
Added missing parametero to raw all workflow
|
2023-02-28 10:16:01 +01:00 |
Claudio Atzori
|
7aebedb43c
|
code formatting
|
2023-02-27 11:51:27 +01:00 |
Miriam Baglioni
|
80987801d7
|
[FoS] added check for null on level1 subject
|
2023-02-27 11:40:22 +01:00 |
Claudio Atzori
|
31e97c2a6b
|
[unresolved entities] updated oozie wf node labels
|
2023-02-27 11:38:29 +01:00 |
Miriam Baglioni
|
23112929e9
|
[FoS] changed the default separator from comma to tab to solve the issue in subject value split
|
2023-02-27 10:18:39 +01:00 |
Serafeim Chatzopoulos
|
0b5bf53b45
|
Remove unecessary indexed fields from Solr
|
2023-02-23 12:42:42 +02:00 |
dimitrispie
|
1547611246
|
Merge branch 'beta' into hive
|
2023-02-22 16:57:12 +02:00 |
Michele Artini
|
fddcf701e9
|
updated the order of the compatibilities
|
2023-02-22 12:07:09 +01:00 |
Claudio Atzori
|
0c1be41b30
|
code formatting
|
2023-02-22 10:15:25 +01:00 |
Claudio Atzori
|
99cd7761aa
|
cleanup of non necessary dhp-monitor-update workflow
|
2023-02-22 10:10:22 +01:00 |
Claudio Atzori
|
cd3a51a15f
|
Merge branch 'beta' into 8232-mdstore-synch-improve
|
2023-02-22 09:57:07 +01:00 |
Claudio Atzori
|
477a7c416f
|
Merge branch 'beta' into UsageCountOnProjectAndDatasource
|
2023-02-22 09:55:51 +01:00 |
Claudio Atzori
|
c20c1c9159
|
Merge pull request 'Added 4 institutions:' (#261) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: #261
|
2023-02-22 09:53:45 +01:00 |
Miriam Baglioni
|
d617c3e812
|
[DOIBoost] extended mapping for funder #8407
|
2023-02-20 14:45:27 +01:00 |
dimitrispie
|
90807b60c7
|
Changes to monitor wf
|
2023-02-20 10:42:24 +02:00 |
dimitrispie
|
d2f9ccf934
|
Changes to separate monitor wf
|
2023-02-20 10:41:21 +02:00 |
dimitrispie
|
032a401cbf
|
Bug fixes
|
2023-02-20 09:29:20 +02:00 |
Miriam Baglioni
|
016337a0f9
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2023-02-16 15:54:59 +01:00 |
Sandro La Bruzzo
|
118c1fc3b3
|
Merge remote-tracking branch 'origin/beta' into beta
|
2023-02-15 10:29:28 +01:00 |
Sandro La Bruzzo
|
a8ac79fa25
|
Added citation relation on crossref Mapping
|
2023-02-15 10:29:13 +01:00 |
dimitrispie
|
595192d510
|
Bug fix
|
2023-02-14 16:24:08 +02:00 |
dimitrispie
|
f3aaff3688
|
Remove duplicate orgs
|
2023-02-14 09:48:36 +02:00 |
Claudio Atzori
|
9a03f71db1
|
code formatting
|
2023-02-13 16:25:47 +01:00 |
Michele Artini
|
554df257ab
|
null values in date range conditions
|
2023-02-13 16:15:32 +01:00 |
dimitrispie
|
3400133c2f
|
Bug fix
|
2023-02-13 09:44:00 +02:00 |
dimitrispie
|
935db0ab25
|
Added organizations for Monitor
|
2023-02-13 09:29:09 +02:00 |
dimitrispie
|
7b78b15c81
|
Changes for copying to Impala Cluster
|
2023-02-13 09:27:00 +02:00 |
Miriam Baglioni
|
5cf902a2b0
|
[UsageCount] changed query to make the sum be computed via sql instead of grouping
|
2023-02-10 16:16:37 +01:00 |
Miriam Baglioni
|
f803530df6
|
[UsageCount] fixed query
|
2023-02-10 15:50:56 +01:00 |
Miriam Baglioni
|
bb5bba51b3
|
[UsageCount] extended test
|
2023-02-09 19:08:30 +01:00 |
Miriam Baglioni
|
85e53fad00
|
[UsageCount] addition of usagecount for Projects and datasources. Extention of the action set created for the results with new entities for projects and datasources. Extention of the resource set and modification of the testing class
|
2023-02-09 18:59:45 +01:00 |
dimitrispie
|
d71f5672d3
|
Add monitor post step
|
2023-02-09 13:44:14 +02:00 |
dimitrispie
|
35ba8bb328
|
Bug fixes
|
2023-02-09 12:57:57 +02:00 |
Sandro La Bruzzo
|
8920932dd8
|
Code formatted
|
2023-02-08 11:34:18 +01:00 |
Sandro La Bruzzo
|
0b9819f1ab
|
Code formatted
|
2023-02-08 10:32:33 +01:00 |
Sandro La Bruzzo
|
6c81a161d2
|
Merge remote-tracking branch 'origin/beta' into 8231-mdstore-synch-improve
|
2023-02-08 10:29:09 +01:00 |
dimitrispie
|
3ba11d64a1
|
Changes 07022023
|
2023-02-07 12:53:51 +02:00 |
dimitrispie
|
98c34263ed
|
Update step20-createMonitorDB.sql
Add University of Cape Town organization
|
2023-02-07 08:14:48 +02:00 |
dimitrispie
|
2dc6d47270
|
Changes 06022023
|
2023-02-06 13:18:53 +02:00 |
dimitrispie
|
973d78a4d6
|
Update step15_5.sql
Added unpaywalls open access colors
|
2023-02-02 08:03:54 +02:00 |
Claudio Atzori
|
d05ca53a14
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2023-01-31 14:39:53 +01:00 |
Miriam Baglioni
|
e82e009b46
|
added missing close tag for XML produced by the xquery to get information for the community from the IS
|
2023-01-31 10:19:34 +01:00 |
Miriam Baglioni
|
b254a0375f
|
[Affiliation from institutionalrepo] changed the field to check to verify the datasource type. Now it is in the field jurisdiction
|
2023-01-26 16:51:20 +01:00 |
dimitrispie
|
cf58e4a5e4
|
Added Arts et Métiers ParisTech
|
2023-01-25 16:03:16 +02:00 |
dimitrispie
|
db7d625ba9
|
Addedd Arts et Métiers ParisTech organization
|
2023-01-25 12:22:21 +02:00 |
Claudio Atzori
|
505867bce9
|
[bulk tagging] better node naming
|
2023-01-20 16:13:16 +01:00 |
Miriam Baglioni
|
ecd398fe51
|
refactoring
|
2023-01-20 14:23:45 +01:00 |
Miriam Baglioni
|
0a5c6010b0
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2023-01-13 16:14:46 +01:00 |
dimitrispie
|
4d7553c9f1
|
Bug fixes
|
2023-01-12 17:19:19 +02:00 |
dimitrispie
|
dd70c32ad7
|
Bug fixes
|
2023-01-12 17:18:05 +02:00 |
dimitrispie
|
51f7ab5864
|
Bug fixes
|
2023-01-12 17:15:06 +02:00 |
dimitrispie
|
34d4bf727c
|
Bug fixes
|
2023-01-12 11:28:37 +02:00 |
dimitrispie
|
43f6d4f296
|
-Monitor DB workflow
|
2023-01-12 11:26:47 +02:00 |
dimitrispie
|
686580a220
|
- New Monitor DB workflow
- New Organization added
|
2023-01-12 11:18:03 +02:00 |
Claudio Atzori
|
0a58bc7ba7
|
[broker] prevent NPEs
|
2023-01-11 14:44:14 +01:00 |
Claudio Atzori
|
04cb96001c
|
[broker] d40e20f437 adapted to the beta graph model
|
2023-01-11 10:10:12 +01:00 |
Michele Artini
|
91b845f611
|
Considering instance pids and alteternative identifiers
|
2023-01-11 09:58:54 +01:00 |
Miriam Baglioni
|
1f367122e4
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2023-01-11 09:47:44 +01:00 |
Michele Artini
|
7b7520850b
|
fixed an invalid char
|
2023-01-11 09:22:18 +01:00 |
Miriam Baglioni
|
d6895f0387
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2023-01-09 17:28:38 +01:00 |
dimitrispie
|
becb242c17
|
Monitor DB only Workflow
|
2023-01-04 16:50:29 +02:00 |
dimitrispie
|
dcb958e146
|
Changes to execute the stats wf only in hive
|
2023-01-04 11:39:01 +02:00 |
dimitrispie
|
592013d5dd
|
Added more steps in decision node
|
2022-12-23 09:43:16 +02:00 |
dimitrispie
|
2a4bf32d4c
|
Merge branch 'hive' of https://code-repo.d4science.org/antonis.lempesis/dnet-hadoop into hive
# Conflicts:
# dhp-workflows/dhp-stats-update/src/main/resources/eu/dnetlib/dhp/oa/graph/stats/oozie_app/scripts/step10.sql
# dhp-workflows/dhp-stats-update/src/main/resources/eu/dnetlib/dhp/oa/graph/stats/oozie_app/scripts/step13.sql
# dhp-workflows/dhp-stats-update/src/main/resources/eu/dnetlib/dhp/oa/graph/stats/oozie_app/scripts/step14.sql
# dhp-workflows/dhp-stats-update/src/main/resources/eu/dnetlib/dhp/oa/graph/stats/oozie_app/scripts/step16_1-definitions.sql
# dhp-workflows/dhp-stats-update/src/main/resources/eu/dnetlib/dhp/oa/graph/stats/oozie_app/scripts/step7.sql
|
2022-12-22 10:22:46 +02:00 |
dimitrispie
|
6449ff4207
|
1. Added a decision node to enables the workflow to make a selection on the execution path to follow
2. Added new organization
3. Added 5 new tables from Eurostast
|
2022-12-22 10:18:21 +02:00 |
Miriam Baglioni
|
8893389895
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2022-12-21 12:42:27 +01:00 |
Antonis Lempesis
|
c8309fe18e
|
addded command line params to allow hive actions to run
|
2022-12-21 12:41:33 +02:00 |
Antonis Lempesis
|
028873cc51
|
added new hive opts
|
2022-12-21 12:41:33 +02:00 |
Antonis Lempesis
|
1ddea4f442
|
removed 'stored as parquet' from views..
|
2022-12-21 12:41:33 +02:00 |
Antonis Lempesis
|
2754c3dd62
|
moving data to impala cluster and creating shadow databases there
|
2022-12-21 12:41:29 +02:00 |
Antonis Lempesis
|
778a1a724f
|
finished migration to hive only
|
2022-12-21 12:41:25 +02:00 |
Antonis Lempesis
|
e84dd5fe26
|
first
|
2022-12-21 12:41:23 +02:00 |
Sandro La Bruzzo
|
3c9826f186
|
updated lines function to it's implementation linesWithSeparators.map(l => l.stripLineEnd) in this way we force scala plugin compiler to consider this pipeline scala code and not java.string.lines() pipeline
|
2022-12-21 11:21:17 +01:00 |
Claudio Atzori
|
6aa91204a5
|
[orcid propagation] skip empty directories
|
2022-12-20 14:15:46 +01:00 |
Miriam Baglioni
|
6674cccb94
|
[BulkTag] description of parameters more comprehensive for those who do not implement it
|
2022-12-16 15:33:20 +01:00 |
Miriam Baglioni
|
f37113a941
|
[BulkTag] moving xquery to get community configuration in dedicated file
|
2022-12-16 15:32:26 +01:00 |
Miriam Baglioni
|
8685eaa706
|
[Clean Country] added test to verify remove of country
|
2022-12-16 15:31:25 +01:00 |
Miriam Baglioni
|
dc0ec88a58
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2022-12-16 13:18:32 +01:00 |
Miriam Baglioni
|
d791840b82
|
[Clean Country] added test to verify remove of country:
|
2022-12-16 13:18:29 +01:00 |
Claudio Atzori
|
7b80b24f82
|
[cleaning] country cleaning must use both PID and AlternateIdentifier fields
|
2022-12-15 14:49:04 +01:00 |
Claudio Atzori
|
b8bafab8a0
|
[cleaning] improved vocabulary based mapping, specialization for the strict vocab cleaning
|
2022-12-12 14:43:03 +01:00 |
Sandro La Bruzzo
|
5e4866d033
|
implemented synch for single mdstore
|
2022-12-12 11:29:46 +01:00 |