miconis
|
fe1c966cbf
|
Merge branch 'master_202203' of code-repo.d4science.org:D-Net/dnet-hadoop into master_202203
|
2022-02-24 17:08:38 +01:00 |
miconis
|
b0f369dc78
|
bug fix in the DedupRecordFactory, DataInfo set before merge
|
2022-02-24 17:08:24 +01:00 |
Miriam Baglioni
|
859cb7ac9d
|
[DoiBoost AR] changed test resource to be sure the result will always have EMBARGO as value for AccessRight
|
2022-02-24 16:55:32 +01:00 |
Miriam Baglioni
|
a40b59b7d5
|
[ResultToOrgFromInstRepoTest] fixed issue in model of the input resources
|
2022-02-24 16:05:57 +01:00 |
Claudio Atzori
|
66c09b1bc7
|
code formatting
|
2022-02-24 12:58:07 +01:00 |
Claudio Atzori
|
a87c070447
|
conflicts resolved, merged from beta
|
2022-02-24 12:51:31 +01:00 |
Claudio Atzori
|
86cdb7a38f
|
[provision] serialize measures defined on the result level
|
2022-02-23 15:54:18 +01:00 |
Alessia Bardi
|
9d6203f79b
|
test mapping datasource
|
2022-02-23 15:00:53 +01:00 |
Antonis Lempesis
|
3b92a2ab9c
|
added the rest of spring 6 in monitor db
|
2022-02-23 12:05:57 +02:00 |
Antonis Lempesis
|
87c91f70a2
|
added sprint 6 indicators to monitor db
|
2022-02-22 14:41:48 +02:00 |
Claudio Atzori
|
5226d0a100
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2022-02-18 15:21:07 +01:00 |
Claudio Atzori
|
99f5b14469
|
[graph raw] invisible records stored among the raw graph rather than the claimed subgraph
|
2022-02-18 15:20:57 +01:00 |
Claudio Atzori
|
401dd38074
|
code formatting
|
2022-02-18 15:19:15 +01:00 |
Claudio Atzori
|
cf8443780e
|
added processingchargeamount to the result view
|
2022-02-18 15:17:48 +01:00 |
Sandro La Bruzzo
|
891781ee3f
|
Merge branch 'beta' of code-repo.d4science.org:D-Net/dnet-hadoop into beta
|
2022-02-18 11:11:32 +01:00 |
Sandro La Bruzzo
|
d3f03abd51
|
fixed wrong json path
|
2022-02-18 11:11:17 +01:00 |
Claudio Atzori
|
89c7313fc5
|
Merge branch 'beta' into hierarchical_orgs_relations
|
2022-02-17 10:30:04 +01:00 |
dimitrispie
|
58c59f46eb
|
Added Sprint 6
|
2022-02-17 10:21:09 +02:00 |
Antonis Lempesis
|
393a4ee956
|
fixed yet another typo...
|
2022-02-15 12:56:50 +02:00 |
Sandro La Bruzzo
|
3aa2020b24
|
added script to regenerate hostedBy Map following instruction defined on ticket #7539
updated hosted By Map
|
2022-02-15 11:05:27 +01:00 |
Miriam Baglioni
|
be64055cfe
|
[OpenCitation] changed the name of destination folders
|
2022-02-14 15:49:44 +01:00 |
Miriam Baglioni
|
1490867cc7
|
[OpenCitation] cleaning of the COCI model
|
2022-02-14 14:52:12 +01:00 |
Miriam Baglioni
|
c191080965
|
mergin with branch beta
|
2022-02-14 14:49:39 +01:00 |
Alessia Bardi
|
600ede1798
|
serialisation of APCs int he XML records
|
2022-02-11 11:00:20 +01:00 |
Miriam Baglioni
|
5c4043dba8
|
[OpenCitation] refactoring
|
2022-02-08 16:23:05 +01:00 |
Miriam Baglioni
|
759ed519f2
|
[OpenCitation] added logic to avoid the genration of self citations relations
|
2022-02-08 16:15:34 +01:00 |
Miriam Baglioni
|
b071f8e415
|
[OpenCitation] change to extract in json format each folder just onece
|
2022-02-08 15:37:28 +01:00 |
Miriam Baglioni
|
fbc28ee8c3
|
[OpenCitation] change the integration logic to consider dois with commas inside
|
2022-02-07 18:32:08 +01:00 |
Miriam Baglioni
|
78be2975f0
|
[stats-wf]fixed another typo related to PR#193
|
2022-02-07 11:22:08 +01:00 |
Miriam Baglioni
|
1f8302dc37
|
Merge pull request '[stats-wf]fixed yet another typo' (#193) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#193
|
2022-02-07 11:19:26 +01:00 |
Antonis Lempesis
|
5f762cbd09
|
fixed yet another typo
|
2022-02-07 12:09:12 +02:00 |
Alessia Bardi
|
ac8b8f224f
|
Merge branch 'beta' into extendResult
|
2022-02-04 16:43:27 +01:00 |
Miriam Baglioni
|
493caef358
|
[stats-wf]fixed the result_result table related to PR#191
|
2022-02-04 14:51:25 +01:00 |
Miriam Baglioni
|
0547fd6ee7
|
Merge pull request '[stats-wf]fixed the result_result table' (#191) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#191
|
2022-02-04 14:47:31 +01:00 |
Antonis Lempesis
|
ae633c566b
|
fixed the result_result table
|
2022-02-04 15:04:19 +02:00 |
Miriam Baglioni
|
aae667e6b6
|
[APC at the result level] added the APC at the level of the result and modified test class
|
2022-02-04 12:34:25 +01:00 |
Sandro La Bruzzo
|
bcfdf9a0d7
|
iis repository with https
|
2022-02-03 16:49:31 +01:00 |
Miriam Baglioni
|
3c60e53a96
|
[stats-wf]fixed the result_result creation for monitor PR#190 on beta
|
2022-02-03 14:47:08 +01:00 |
Miriam Baglioni
|
89922156c9
|
Merge pull request '[stats-wf]fixed the result_result creation for monitor' (#190) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#190
|
2022-02-03 13:00:56 +01:00 |
Antonis Lempesis
|
c2b44530a3
|
typo...
|
2022-02-03 13:44:07 +02:00 |
Antonis Lempesis
|
dbd2646d59
|
fixed the result_result creation for monitor
|
2022-02-03 12:37:10 +02:00 |
Alessia Bardi
|
2e215abfa8
|
test for instances with URLs for OpenAPC
|
2022-02-02 17:27:44 +01:00 |
Miriam Baglioni
|
37784209c9
|
[dhp-schemas-] updated the version of dhp-schema to 2.10.27 for APC name and id modification
|
2022-02-02 12:46:31 +01:00 |
Miriam Baglioni
|
73eba34d42
|
[UnresolvedEntities] Changed the way to merge the unresolved because the new merge removed the dataInfo from the merged result. Added also data info for subjects
|
2022-02-01 08:38:41 +01:00 |
Miriam Baglioni
|
dce7f5fea8
|
[BULK TAGGING] changed to fix issue that should have been fixed already
|
2022-01-31 08:20:28 +01:00 |
Claudio Atzori
|
8eb75ca169
|
adapted GenerateEntitiesApplicationTest behaviour
|
2022-01-27 16:24:37 +01:00 |
Claudio Atzori
|
af61e44acc
|
ported changes to the GraphCleaningFunctionsTest from 8de9788308
|
2022-01-27 16:19:14 +01:00 |
Claudio Atzori
|
1322379741
|
Merge branch 'beta' into delegated_authorities
|
2022-01-25 14:28:25 +01:00 |
Claudio Atzori
|
59a250337c
|
[graph resolution] drop output path at the beginning
|
2022-01-24 18:02:39 +01:00 |
Claudio Atzori
|
97ad94d7d9
|
[graph resolution] drop output path at the beginning
|
2022-01-24 18:02:07 +01:00 |
Claudio Atzori
|
8de9788308
|
applied fix for avoiding ruling out the invisible (APC) records during the graph cleaning
|
2022-01-24 11:29:22 +01:00 |
Claudio Atzori
|
2f385b3ac6
|
updated dnet workflow profile definitions
|
2022-01-21 13:59:46 +01:00 |
Claudio Atzori
|
dd52bf1bb8
|
copy relations to the graphOutputPath
|
2022-01-21 13:59:29 +01:00 |
Claudio Atzori
|
4983d6536d
|
Merge branch 'beta' into delegated_authorities
|
2022-01-21 13:02:48 +01:00 |
Claudio Atzori
|
f0ea2410e5
|
improved mapping titles from datacite records to consider title types
|
2022-01-21 10:50:34 +01:00 |
Claudio Atzori
|
b37bc277c4
|
reintroduced the hostedby patching to the datacite records
|
2022-01-21 09:15:13 +01:00 |
Claudio Atzori
|
f2fde5566b
|
using helper method from ModelSupport to find the inverse relation descriptor
|
2022-01-20 09:19:07 +01:00 |
Claudio Atzori
|
3b9020c1b7
|
added unit test for the DispatchEntitiesJob
|
2022-01-19 18:15:55 +01:00 |
Claudio Atzori
|
abfa9c6045
|
code formatting
|
2022-01-19 17:17:11 +01:00 |
Claudio Atzori
|
391aa1373b
|
added unit test
|
2022-01-19 17:13:21 +01:00 |
Claudio Atzori
|
44a937f4ed
|
factored out entity grouping implementation, extended to consider results from delegated authorities rather than identical records from other sources
|
2022-01-19 12:24:52 +01:00 |
Miriam Baglioni
|
a7c4d0d16d
|
[DoiBoost Organizations] added parameter to specify the action in the wf raw_organizations to be able to load the openorgs organization as in the loading step for the construction of the graph
|
2022-01-13 13:52:00 +01:00 |
Miriam Baglioni
|
a75fb8c47a
|
[BipFinderInstanceLevel] change pom to align to the dhp-schema release 2.10.24 and refactoring
|
2022-01-12 18:06:26 +01:00 |
Miriam Baglioni
|
e7d5a39c03
|
[BipFinderInstanceLevel] added tests in test class
|
2022-01-12 17:25:04 +01:00 |
Miriam Baglioni
|
4993666d73
|
[BipFinderInstanceLevel] changed creation of the instance to allow to enrich existing instances with same pid
|
2022-01-12 16:53:47 +01:00 |
Claudio Atzori
|
9acc32faa6
|
[stats wf] final touches for the integration of PRs #166, #179 in the master branch
|
2022-01-12 12:04:31 +01:00 |
dimitrispie
|
b053b0178e
|
Sprint 5 and other changes
|
2022-01-12 11:23:37 +01:00 |
Antonis Lempesis
|
b6b4bc0df9
|
added first indicator of sprint 5
|
2022-01-12 11:20:28 +01:00 |
Antonis Lempesis
|
e91f06f39b
|
fixed typos in indicators. Added extra views in monitor
|
2022-01-12 11:18:28 +01:00 |
Antonis Lempesis
|
3ce1976627
|
fixed column names
|
2022-01-12 11:14:41 +01:00 |
Antonis Lempesis
|
4878d7485c
|
added usage stats
|
2022-01-12 11:13:25 +01:00 |
Antonis Lempesis
|
a4316bafed
|
fixed a typo
|
2022-01-12 11:12:53 +01:00 |
Antonis Lempesis
|
bb17e070d8
|
added result_result relations
|
2022-01-12 11:09:38 +01:00 |
Claudio Atzori
|
a30a98a716
|
Applying PR#166 in the master branch (Added sprint 3&4 of indicators). Merge commit '0df9574a6f5d9d75bc840decb023561ae941f9d6'
|
2022-01-12 10:57:19 +01:00 |
Sandro La Bruzzo
|
57e2c4b749
|
formatted code
|
2022-01-12 09:40:28 +01:00 |
Claudio Atzori
|
0f2144b5e0
|
scalafmt: code formatting
|
2022-01-11 17:03:44 +01:00 |
Claudio Atzori
|
dcd282977c
|
pulled from beta
|
2022-01-11 16:59:41 +01:00 |
Claudio Atzori
|
4f212652ca
|
scalafmt: code formatting
|
2022-01-11 16:57:48 +01:00 |
Sandro La Bruzzo
|
0163dadb7f
|
[doiboost]
- update MAG schema, new filed added on version dec-2021
|
2022-01-11 11:05:44 +01:00 |
Miriam Baglioni
|
904e1c2667
|
Merge pull request 'Affiliation Propagation through semantic relation' (#183) from enrichment into beta
Reviewed-on: D-Net/dnet-hadoop#183
|
2022-01-07 19:18:16 +01:00 |
Miriam Baglioni
|
064f9bbd87
|
[AFFPropSR] added new paprameter for the number of iterations and new code for just one iteration
|
2022-01-07 18:58:51 +01:00 |
Miriam Baglioni
|
b7e450070b
|
[SDG-FOS] to import SDG file not considering the header
|
2022-01-07 12:13:26 +01:00 |
Miriam Baglioni
|
639190370a
|
mergin with branch beta
|
2022-01-07 11:29:25 +01:00 |
Miriam Baglioni
|
adccc2346a
|
[SDG-FOS] to lower case for the doi
|
2022-01-07 11:28:50 +01:00 |
Claudio Atzori
|
8ae46ca789
|
OAF-store-graph mdstores: firther fix for PR#180
|
2022-01-05 15:52:15 +01:00 |
Claudio Atzori
|
908294d86e
|
OAF-store-graph mdstores: firther fix for PR#180
|
2022-01-05 15:49:05 +01:00 |
Claudio Atzori
|
3bd3653be9
|
OAF-store-graph mdstores: save them in text format
|
2022-01-04 16:39:39 +01:00 |
Claudio Atzori
|
3dc48c7ab5
|
OAF-store-graph mdstores: save them in text format
|
2022-01-04 16:39:27 +01:00 |
Claudio Atzori
|
f82db765db
|
OAF-store-graph mdstores: save them in text format
|
2022-01-04 16:39:15 +01:00 |
Claudio Atzori
|
8d13effa31
|
test for the tolerant deserialisation utility method
|
2022-01-04 16:38:26 +01:00 |
Claudio Atzori
|
9458ee7938
|
serialise records in the OAF-store-graph mdstores in json format. Read them again in the graph construction phase using a tolerant parser to support backward compatible changes in the evolution of the schema
|
2022-01-04 16:38:09 +01:00 |
Claudio Atzori
|
58f8998e3d
|
OAF-store-graph mdstores: save them in text format
|
2022-01-04 15:02:09 +01:00 |
Claudio Atzori
|
174c3037e1
|
OAF-store-graph mdstores: save them in text format
|
2022-01-04 14:40:16 +01:00 |
Claudio Atzori
|
045d767013
|
OAF-store-graph mdstores: save them in text format
|
2022-01-04 14:23:01 +01:00 |
Claudio Atzori
|
bd59b58efb
|
test for the tolerant deserialisation utility method
|
2022-01-04 11:26:56 +01:00 |
Claudio Atzori
|
a6977197b3
|
serialise records in the OAF-store-graph mdstores in json format. Read them again in the graph construction phase using a tolerant parser to support backward compatible changes in the evolution of the schema
|
2022-01-03 17:25:26 +01:00 |
Miriam Baglioni
|
4c60ee1718
|
mergin with branch beta
|
2022-01-03 15:24:02 +01:00 |
Miriam Baglioni
|
92fd69e25d
|
[SDG-FOS] alternative way to get input data to avoid OOM error while getting csv
|
2022-01-03 15:23:06 +01:00 |
Claudio Atzori
|
fe7e5f4748
|
Merge pull request '[stats wf] result_result relations, usage stats, monitor views, indicator for sprint 5' (#179) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#179
|
2022-01-03 14:52:11 +01:00 |
Claudio Atzori
|
bcea4e3a9b
|
added dnet workflow profile for the orchestration of the simplified and complete graph construction and processing pipeline, where the IIS works on the non-deduplicated graph
|
2022-01-03 14:33:00 +01:00 |
Miriam Baglioni
|
a706ba0c08
|
Merge pull request 'SDG Integration' (#178) from SDG into beta
Reviewed-on: D-Net/dnet-hadoop#178
|
2021-12-23 14:50:00 +01:00 |
Antonis Lempesis
|
81ee654271
|
added result_result relations
|
2021-12-23 15:46:17 +02:00 |
Antonis Lempesis
|
7551e52e95
|
fixed a typo
|
2021-12-23 15:33:53 +02:00 |
Miriam Baglioni
|
7a1b440413
|
[SDG] logic to create unresolved entities out of SDG input. This changes also some classes related to FOS to reuse the same code. The code under createunresolvedentities create results with the merged update of the the inputs provided (bip at the level of the isntance, fos and sdg for subjects)
|
2021-12-23 13:24:28 +01:00 |
Claudio Atzori
|
cccb16900c
|
https://support.openaire.eu/issues/7330 normalising DOI urls
|
2021-12-23 12:33:53 +01:00 |
Miriam Baglioni
|
2a67ee13ec
|
[SDG] added model class
|
2021-12-23 10:37:52 +01:00 |
Miriam Baglioni
|
69e9ea9eeb
|
[Graph Dump] Test for extraction of rels from entities extended
|
2021-12-23 10:15:30 +01:00 |
Miriam Baglioni
|
31b26d48ac
|
[Graph Dump] fixed issue on extraction of relation between entities and contexts: the relationship name and type were swapped
|
2021-12-23 10:09:47 +01:00 |
Miriam Baglioni
|
10579c0dd0
|
[FOS]fixed doi value in test
|
2021-12-22 23:10:16 +01:00 |
Miriam Baglioni
|
6116fc5d40
|
[FOS]added logic to include only different subjects. Test refactoring and extention
|
2021-12-22 23:04:22 +01:00 |
Miriam Baglioni
|
b81efb6a9d
|
[FOS]changed the mapping between the csv and the model. Changed Test classes and resources
|
2021-12-22 21:40:35 +01:00 |
Miriam Baglioni
|
de6c4c8968
|
[FOS]creation of the unresolved entities: remove the split for the doi: no more needed since each row is related to one doi
|
2021-12-22 16:44:44 +01:00 |
Miriam Baglioni
|
34ac56565d
|
refactoring
|
2021-12-22 16:28:11 +01:00 |
Miriam Baglioni
|
20ef1d657f
|
refactoring
|
2021-12-22 16:26:36 +01:00 |
Miriam Baglioni
|
813f856d3f
|
[BipFinder] removing left over parameter in wf
|
2021-12-22 16:11:12 +01:00 |
Miriam Baglioni
|
2c126ed014
|
[BipFinder] create unresolved entities with measures at the level of the instance
|
2021-12-22 16:03:41 +01:00 |
Miriam Baglioni
|
0807fdb65a
|
[BipFinder] remove not needed resources
|
2021-12-22 15:37:00 +01:00 |
Miriam Baglioni
|
b5e11a3a0a
|
[BipFinder] put in common package BipFinder model
|
2021-12-22 15:33:05 +01:00 |
Miriam Baglioni
|
c5739c4266
|
[BipFinder] create action set for the measures at the level of the result
|
2021-12-22 15:08:33 +01:00 |
Miriam Baglioni
|
da5f6260aa
|
mergin with branch beta
|
2021-12-22 13:12:02 +01:00 |
Miriam Baglioni
|
be0acccf42
|
Merge branch 'beta' into dump
|
2021-12-22 12:39:57 +01:00 |
Antonis Lempesis
|
16539d7360
|
added usage stats
|
2021-12-22 02:54:42 +02:00 |
Antonis Lempesis
|
3edd661608
|
fixed column names
|
2021-12-21 22:55:04 +02:00 |
Antonis Lempesis
|
a4c0cbb98c
|
fixed typos in indicators. Added extra views in monitor
|
2021-12-21 15:54:38 +02:00 |
Miriam Baglioni
|
e24a7f3496
|
mergin with branch beta
|
2021-12-21 13:57:19 +01:00 |
Miriam Baglioni
|
d1ae219cb4
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-12-21 13:55:53 +01:00 |
Miriam Baglioni
|
460e6b95d6
|
[Graph Dump] -
|
2021-12-21 13:48:03 +01:00 |
Sandro La Bruzzo
|
3920d68992
|
Fixed workflow generation of delta in datacite
|
2021-12-21 11:41:49 +01:00 |
Antonis Lempesis
|
58996972d9
|
added first indicator of sprint 5
|
2021-12-21 03:35:04 +02:00 |
dimitrispie
|
c1cdec09a9
|
Sprint 5 and other changes
|
2021-12-20 19:23:57 +02:00 |
Miriam Baglioni
|
3cc1b7b153
|
mergin with branch beta
|
2021-12-15 17:25:02 +01:00 |
Miriam Baglioni
|
63b648b0dd
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-12-15 12:41:15 +01:00 |
Antonis Lempesis
|
f0b523cfa7
|
removed the too restrctive clause. will discuss again
|
2021-12-15 12:32:15 +01:00 |
Sandro La Bruzzo
|
b881ee5ef8
|
[scholexplorer]
- implemented generation of scholix of delta update of datacite
|
2021-12-15 11:25:32 +01:00 |
Sandro La Bruzzo
|
63952018c0
|
[scholexplorer]
-moved SparkRetrieveDataciteDelta in scala folder
|
2021-12-15 11:25:32 +01:00 |
Sandro La Bruzzo
|
e5bff64f2e
|
[scholexplorer]
- Minor fix on SparkConvertRDDtoDataset
-first implementation of retrieve datacite dump
|
2021-12-15 11:25:32 +01:00 |
Claudio Atzori
|
1790fa2d44
|
Merge branch 'beta' into affiliationPropagation
|
2021-12-14 15:26:56 +01:00 |
Miriam Baglioni
|
56409d1281
|
[Dump] resolved conflicts with beta and merging
|
2021-12-14 15:03:45 +01:00 |
Miriam Baglioni
|
22d4b5619b
|
[BipFinder Result] last changes to test and resources files
|
2021-12-14 14:54:13 +01:00 |
Miriam Baglioni
|
6fb6236cd4
|
changed the way to produce the AS for bipFinder.
|
2021-12-14 14:51:14 +01:00 |
Miriam Baglioni
|
573bd17cbb
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-12-14 11:12:25 +01:00 |
Miriam Baglioni
|
4eb8276493
|
-
|
2021-12-14 11:12:17 +01:00 |
Miriam Baglioni
|
936578aaf1
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-12-13 15:01:47 +01:00 |
Miriam Baglioni
|
8d755cca80
|
-
|
2021-12-13 15:01:40 +01:00 |
Claudio Atzori
|
98eb292c59
|
avoid NPEs merging XMLInstance(s)
|
2021-12-13 13:27:20 +01:00 |
Claudio Atzori
|
5e17247bb6
|
avoid NPEs merging XMLInstance(s)
|
2021-12-13 11:48:40 +01:00 |
Claudio Atzori
|
b70ecccea0
|
avoid NPEs merging XMLInstance(s)
|
2021-12-12 12:37:38 +01:00 |
Claudio Atzori
|
c1b6ae47cd
|
cleaning workflow assigns the proper default instance type when a value could not be cleaned using the vocabularies
|
2021-12-09 16:47:41 +01:00 |
Claudio Atzori
|
eb43eda42a
|
Merge branch 'beta' into graph_cleaning
|
2021-12-09 16:46:48 +01:00 |
Claudio Atzori
|
41c70c607d
|
cleaning workflow assigns the proper default instance type when a value could not be cleaned using the vocabularies
|
2021-12-09 16:44:28 +01:00 |
Alessia Bardi
|
cba63e9f82
|
Merge branch 'beta' into sygma_indexing
|
2021-12-09 15:52:16 +01:00 |
Alessia Bardi
|
e53228401b
|
style
|
2021-12-09 15:46:22 +01:00 |
Claudio Atzori
|
cd9c51fd7a
|
vocabulary based cleaning considers also the term label when looking up for a synonym
|
2021-12-09 14:49:24 +01:00 |
Claudio Atzori
|
e6e177dda0
|
vocabulary based cleaning considers also the term label when looking up for a synonym
|
2021-12-09 13:57:53 +01:00 |
Alessia Bardi
|
6b5d7688a4
|
#7275 serialize license information in XML records
|
2021-12-09 13:46:48 +01:00 |
Miriam Baglioni
|
b113586207
|
resolved conflicts
|
2021-12-07 10:16:14 +01:00 |
Sandro La Bruzzo
|
5d51b3dd4a
|
Merge pull request 'scala_refactor' (#169) from scala_refactor into beta
Reviewed-on: D-Net/dnet-hadoop#169
|
2021-12-06 15:33:44 +01:00 |
Miriam Baglioni
|
d9836f0cf3
|
[OpenCitations] fixed test when executed one after the other
|
2021-12-06 15:27:09 +01:00 |
Miriam Baglioni
|
d1df01ff1e
|
[Graph Dump] fixed resource for test
|
2021-12-06 15:15:48 +01:00 |
Sandro La Bruzzo
|
ed0c352799
|
[test-fixing] fixed wrong test
|
2021-12-06 15:07:41 +01:00 |
Miriam Baglioni
|
96a7d46278
|
[Graph Dump] fixed tests
|
2021-12-06 15:06:32 +01:00 |
Sandro La Bruzzo
|
e9f285ec4d
|
[scala-refactor] Module dhp-doiboost:
Moved all scala source into src/main/scala and src/test/scala
|
2021-12-06 14:24:03 +01:00 |
Sandro La Bruzzo
|
bf880e2508
|
[scala-refactor] Module dhp-graph-mapper:
Moved all scala source into src/main/scala and src/test/scala
|
2021-12-06 13:57:41 +01:00 |
Sandro La Bruzzo
|
7af0bbd0b1
|
[scala-refactor] Module dhp-aggregation:
Moved all scala source into src/main/scala and src/test/scala
|
2021-12-06 11:26:36 +01:00 |
Claudio Atzori
|
08795cbd30
|
using helper method from ModelSupport to find the inverse relation descriptor
|
2021-12-06 10:39:56 +01:00 |
Miriam Baglioni
|
f430688ff7
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-12-03 12:36:08 +01:00 |
Miriam Baglioni
|
4bb1d43afc
|
-
|
2021-12-03 12:35:51 +01:00 |
Sandro La Bruzzo
|
f7011b90d8
|
format code
|
2021-12-03 11:15:09 +01:00 |
Claudio Atzori
|
dd0b2e5244
|
Merge branch 'beta' into instance_group_by_url
|
2021-12-03 09:27:58 +01:00 |
Claudio Atzori
|
863a2f9db3
|
avoid to filter OAF records defined as invisible = true
|
2021-12-03 09:08:12 +01:00 |
Claudio Atzori
|
9cac283bec
|
implemented Instance serialization features requested in https://support.openaire.eu/issues/7156
|
2021-12-02 17:20:33 +01:00 |
Miriam Baglioni
|
d9f80488cc
|
[GRAPH DUMP] Add one more test to check the filtering of the relations
|
2021-12-02 14:15:19 +01:00 |
Miriam Baglioni
|
58bc3f223a
|
[GRAPH DUMP] Add filtering for relation we do not want to dump. It is based on the relclass
|
2021-12-02 14:09:46 +01:00 |
Miriam Baglioni
|
8905a39bf3
|
mergin with branch beta
|
2021-12-02 13:17:29 +01:00 |
Miriam Baglioni
|
87eedad898
|
-
|
2021-12-02 13:17:19 +01:00 |
Claudio Atzori
|
3b19821f3c
|
added stats computation on the graph hive DB tables
|
2021-12-02 10:44:10 +01:00 |
Claudio Atzori
|
cfa4560769
|
minor: fixed hive action name
|
2021-12-02 10:43:36 +01:00 |
Claudio Atzori
|
d85af6fc25
|
[cleaning wf] fixed OAF record navigation, a mapping defined on a container object would have prevented the natvigation to continue on its properties
|
2021-12-01 15:49:15 +01:00 |
Claudio Atzori
|
4fe7888817
|
code formatting
|
2021-12-01 15:48:15 +01:00 |
Claudio Atzori
|
01e5e0142a
|
added test to verify the relation inverse lookup operation
|
2021-12-01 09:46:26 +01:00 |
Claudio Atzori
|
0df9574a6f
|
Merge pull request '[stats wf] Added sprint 3&4 of indicators' (#166) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#166
|
2021-11-29 10:40:26 +01:00 |
Claudio Atzori
|
1de881b796
|
resolved conflicts for #165
|
2021-11-26 16:15:11 +01:00 |
Claudio Atzori
|
014e872ae1
|
[resolution wf] added optional parameter to skip the entity resolution
|
2021-11-26 15:38:56 +01:00 |
Claudio Atzori
|
5c6d328537
|
code formatting
|
2021-11-26 15:38:16 +01:00 |
dimitrispie
|
09fc2afdca
|
Added indi_funder_country_collab
Kept only indi_pub_has_cc_licence
|
2021-11-26 16:13:10 +02:00 |
Antonis Lempesis
|
0b4163ee0b
|
added sprint3,4, removed 2, chaos
|
2021-11-26 15:58:01 +02:00 |
dimitrispie
|
29f69f2f89
|
Sprint 4
|
2021-11-26 15:22:04 +02:00 |
Miriam Baglioni
|
ac07ed8251
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-11-25 12:32:58 +01:00 |
Miriam Baglioni
|
5fd0e610bf
|
[DOIBOOST Process] fix filtering to filter results with non null id
|
2021-11-25 12:10:45 +01:00 |
Sandro La Bruzzo
|
feea154e89
|
remove working dir after test
|
2021-11-25 11:02:38 +01:00 |
Sandro La Bruzzo
|
028a8acad8
|
add test resources
|
2021-11-25 10:54:47 +01:00 |
Sandro La Bruzzo
|
2164a2a889
|
Datacite: Code Refactor generated a general SparkApplication Scala where all the spark scala have to inherit
Commented a little the Datacite transformation code
|
2021-11-25 10:54:13 +01:00 |
Miriam Baglioni
|
3f9b2ba8ce
|
[Hosted By Map] fix issue in test
|
2021-11-22 16:59:43 +01:00 |
Sandro La Bruzzo
|
a7cf277d98
|
Datacite: Removed HostedBy Patch as described on ticket #7219, Now all the records will have hosted by Unknown Repository
|
2021-11-22 16:03:17 +01:00 |
Sandro La Bruzzo
|
483d3039d1
|
entity resolution: added distcpt of missing entities in graph materialization
|
2021-11-22 15:55:24 +01:00 |
Sandro La Bruzzo
|
93fe8ce8b2
|
entity resolution: fix test
|
2021-11-22 15:50:43 +01:00 |
Sandro La Bruzzo
|
35e20b0647
|
updated resolution wf:
- generate a new version of the graph
- changed merge from union to join
|
2021-11-22 11:48:55 +01:00 |
Miriam Baglioni
|
fdb75b180e
|
[Cleaning] added couple of tests for DOIBOOST publications
|
2021-11-21 16:35:22 +01:00 |
Miriam Baglioni
|
0506fa2654
|
[Graph Dump] changed to mirror the changes in the model
|
2021-11-19 15:56:25 +01:00 |
Sandro La Bruzzo
|
3426451d3f
|
Merge remote-tracking branch 'origin/beta' into beta
|
2021-11-19 14:49:04 +01:00 |
Sandro La Bruzzo
|
4542a2338b
|
updated site configuration to deploy on website
|
2021-11-19 13:44:08 +01:00 |
Claudio Atzori
|
e5a2c596b2
|
Merge branch 'beta' into preserve_openorg_parent_child_relations
|
2021-11-19 11:35:46 +01:00 |
Claudio Atzori
|
f4538f3c4c
|
cleanup
|
2021-11-19 11:33:10 +01:00 |
Claudio Atzori
|
2b46b87f56
|
fixed filtering criteria applied in SparkCopyRelationsNoOpenorgs to keep the parent/child relations from OpenOrgs
|
2021-11-19 11:30:29 +01:00 |
Miriam Baglioni
|
9fae872181
|
[Graph Dump] changed to mirror the changes in the model
|
2021-11-19 11:25:50 +01:00 |
Sandro La Bruzzo
|
fc03c99805
|
fixed javadocs url after deploying site
|
2021-11-19 10:46:33 +01:00 |
Sandro La Bruzzo
|
0c0d561bc4
|
added public class into tests to create correct javadoc
|
2021-11-19 09:54:22 +01:00 |
Claudio Atzori
|
62fa61f3cf
|
merge from beta
|
2021-11-19 09:23:42 +01:00 |
Claudio Atzori
|
bd9a43cefd
|
Revert to 4094f2bb9a
|
2021-11-19 09:20:43 +01:00 |
Claudio Atzori
|
3a4d925386
|
Merge branch 'beta' into hierarchical_orgs_relations
|
2021-11-18 18:07:08 +01:00 |
Claudio Atzori
|
3974fa7dc1
|
Merge branch 'beta' into affiliationPropagation
|
2021-11-18 18:06:26 +01:00 |
Claudio Atzori
|
a24b9f8268
|
[dedup] trivial refactoring
|
2021-11-18 17:12:02 +01:00 |
Claudio Atzori
|
c0750fb17c
|
avoid non necessary count operations over large spark datasets
|
2021-11-18 17:11:31 +01:00 |
Claudio Atzori
|
bb5dca7979
|
cleanup
|
2021-11-18 17:10:46 +01:00 |
Miriam Baglioni
|
793b5a8e5f
|
Aggiornare 'dhp-workflows/dhp-graph-mapper/src/main/java/eu/dnetlib/dhp/oa/graph/dump/ResultMapper.java'
Removing the dump of Measure at the level of the result. We decided not to map it
|
2021-11-18 14:49:38 +01:00 |
Miriam Baglioni
|
5dc5792722
|
[Graph Dump] Change test resource to mirror the movement of the measure element
|
2021-11-18 14:39:12 +01:00 |
Miriam Baglioni
|
0136a8c266
|
[Graph Dump] Change test to mirror that measure is at the level of the isntance
|
2021-11-18 14:38:33 +01:00 |
Miriam Baglioni
|
1b79c0ee79
|
mergin with branch beta
|
2021-11-18 11:01:00 +01:00 |
Antonis Lempesis
|
cb3adb90f4
|
Merge branch 'beta' into beta
|
2021-11-17 14:33:45 +01:00 |
Antonis Lempesis
|
c283406829
|
added Universidad Polytecnica de Madrid
|
2021-11-17 15:33:00 +02:00 |
Claudio Atzori
|
e0395719d7
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-11-17 14:17:27 +01:00 |
Claudio Atzori
|
82a4e4efae
|
[cleaning wf] fixed methodology to rule out invalid result titles, based on https://support.openaire.eu/issues/7206
|
2021-11-17 14:17:22 +01:00 |
Miriam Baglioni
|
6d4a1c57ee
|
[Resolve Entities] Change test dataset to mirror the modification in the creation of the map between the pids and the unresolved
|
2021-11-17 12:41:52 +01:00 |
Sandro La Bruzzo
|
9c82d670b8
|
make class public in order to create javadoc
|
2021-11-17 12:31:02 +01:00 |
Sandro La Bruzzo
|
1f5ee116ed
|
code refactor, created and moved scala code on the correct maven folder under src/main/scala and src/test/scala
fixed test
|
2021-11-17 12:23:52 +01:00 |
Sandro La Bruzzo
|
2fd9ceac13
|
code refactor, created and moved scala code on the correct maven folder under src/main/scala and src/test/scala
|
2021-11-17 11:35:22 +01:00 |
Sandro La Bruzzo
|
2506d7a679
|
Merge branch 'mvn_site_documentation' of code-repo.d4science.org:D-Net/dnet-hadoop into mvn_site_documentation
|
2021-11-17 11:07:24 +01:00 |
Sandro La Bruzzo
|
cded363b55
|
code refactor, created and moved scala code on the correct maven folder under src/main/scala and src/test/scala
|
2021-11-17 11:06:35 +01:00 |
Miriam Baglioni
|
4094f2bb9a
|
added integration md file
|
2021-11-17 10:04:52 +01:00 |
Miriam Baglioni
|
ec8b0219ff
|
[Documentation] Added first page for Integration via unresolved entities generation
|
2021-11-16 17:41:34 +01:00 |
Miriam Baglioni
|
2bbece2ca5
|
mergin with branch beta
|
2021-11-16 16:35:40 +01:00 |
Sandro La Bruzzo
|
2d67020c59
|
added dhp-enrichment maven site template
|
2021-11-16 16:01:08 +01:00 |
Miriam Baglioni
|
28ea532ece
|
[Affilaition Propagation] moved the selection of graph relation as a preparation step
|
2021-11-16 15:24:19 +01:00 |
Sandro La Bruzzo
|
18c1d70ef4
|
Merge branch 'beta' of code-repo.d4science.org:D-Net/dnet-hadoop into mvn_site_documentation
|
2021-11-16 15:16:49 +01:00 |
Sandro La Bruzzo
|
a1cafaf2e3
|
added mvn site for dnet-hadoop project
|
2021-11-16 15:16:28 +01:00 |
Miriam Baglioni
|
7c96e3fd46
|
removed not useful dir
|
2021-11-16 13:57:26 +01:00 |
Miriam Baglioni
|
c7c0c3187b
|
[AFFILIATION PROPAGATION] Applied some SonarLint suggestions
|
2021-11-16 13:56:32 +01:00 |
Miriam Baglioni
|
c6a9f0a1a8
|
mergin with branch beta
|
2021-11-16 12:04:40 +01:00 |
Miriam Baglioni
|
99d86134f5
|
[Graph Dump] changed the dump since the measures have been moded at the level of the instance
|
2021-11-16 12:04:21 +01:00 |
Claudio Atzori
|
0a727d325d
|
[dedup] increased number of partitions in the consistency phase
|
2021-11-16 08:43:41 +01:00 |
Claudio Atzori
|
bafa2990f3
|
code formatting
|
2021-11-15 17:07:16 +01:00 |
Claudio Atzori
|
668ac25224
|
[graph resolution] using existing argument parser file name
|
2021-11-15 17:02:45 +01:00 |
Claudio Atzori
|
7d0a03f607
|
[graph resolution] minor
|
2021-11-15 14:45:54 +01:00 |
Claudio Atzori
|
941a50a2fc
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2021-11-15 14:42:49 +01:00 |
Claudio Atzori
|
7c804acda8
|
[graph resolution] minor
|
2021-11-15 14:42:43 +01:00 |
Sandro La Bruzzo
|
efa09057db
|
Merge branch 'beta' of code-repo.d4science.org:D-Net/dnet-hadoop into beta
|
2021-11-15 14:32:09 +01:00 |
Sandro La Bruzzo
|
48923e46a1
|
added documentation to Pubmed Class and also added mvn site for dhp-aggregations
|
2021-11-15 14:32:01 +01:00 |
Claudio Atzori
|
d2c787d416
|
[graph resolution] fixed sequence of the workflow steps
|
2021-11-15 14:31:15 +01:00 |
Claudio Atzori
|
975b10b711
|
[actionmanager] increased spark.sql.shuffle.partitions to 5000
|
2021-11-15 12:31:45 +01:00 |
Miriam Baglioni
|
4ec88c718c
|
merge with beta - resolved conflict in pom
|
2021-11-15 10:52:16 +01:00 |
Miriam Baglioni
|
6f1a434e90
|
[Bypass Action Set] Fixed test to consider the new identifier utils
|
2021-11-15 09:59:23 +01:00 |
Miriam Baglioni
|
157d33ebf9
|
[Bypass Action Set] Refactoring
|
2021-11-15 09:58:48 +01:00 |
Miriam Baglioni
|
6595135a1a
|
[Dump Schemas] changed the schema of the dumped result according to the modifications in the bestAccessRight type
|
2021-11-12 11:45:38 +01:00 |
Miriam Baglioni
|
43cae4ad88
|
Merge branch 'dump' of https://code-repo.d4science.org/D-Net/dnet-hadoop into dump
|
2021-11-12 11:36:54 +01:00 |
Miriam Baglioni
|
b3f9370125
|
merge with beta - resolved conflict in pom
|
2021-11-12 11:25:26 +01:00 |
Miriam Baglioni
|
92d0e18b55
|
[Bypass Action Set] used constant DOI instead of "doi"
|
2021-11-12 10:56:58 +01:00 |
Miriam Baglioni
|
881113743f
|
[Bypass Action Set] refactoring
|
2021-11-12 10:55:50 +01:00 |
Miriam Baglioni
|
47ccb53c4f
|
[Bypass Action Set] modification for comment D-Net/dnet-hadoop#157 (comment)
|
2021-11-12 10:54:09 +01:00 |
Miriam Baglioni
|
ffb0ce1d59
|
merge with beta - resolved conflict in pom
|
2021-11-12 10:19:59 +01:00 |
Miriam Baglioni
|
716021546e
|
[Bypass Action Set] minor fix
|
2021-11-12 10:18:01 +01:00 |
Sandro La Bruzzo
|
3469cc2b1d
|
Merge branch 'beta' of code-repo.d4science.org:D-Net/dnet-hadoop into beta
|
2021-11-12 09:56:52 +01:00 |
Sandro La Bruzzo
|
a7763d2492
|
removed alternate identifier in resolutionMap
|
2021-11-12 09:56:45 +01:00 |
Miriam Baglioni
|
b8bdabfae9
|
[Graph DUmp] removed OpenAccessRoute from test in best access right
|
2021-11-11 16:16:48 +01:00 |
Miriam Baglioni
|
e5498052e8
|
[Graph DUmp] removed OpenAccessRoute from test in best access right
|
2021-11-11 16:14:10 +01:00 |
Miriam Baglioni
|
935062edec
|
[Bypass Action Set] creation of unresolved entities
|
2021-11-11 16:11:25 +01:00 |
Antonis Lempesis
|
26f086dd64
|
removed the too restrctive clause. will discuss again
|
2021-11-11 12:57:19 +02:00 |
Claudio Atzori
|
148289150f
|
Merge branch 'beta' into doiboost_url
|
2021-11-11 10:40:19 +01:00 |
Sandro La Bruzzo
|
2ca0a436ad
|
added SparkResolveEntities node to the oozie wf
|
2021-11-11 10:25:42 +01:00 |
Sandro La Bruzzo
|
9cb195314f
|
implemented and tested resolution of entities
|
2021-11-11 10:17:40 +01:00 |
Miriam Baglioni
|
6d3c4c4abe
|
mergin with branch beta
|
2021-11-11 08:59:53 +01:00 |
Miriam Baglioni
|
8cc50ecee0
|
[Graph Dump] changed AccessRight with BestAccessRight in the dump and modified the dependency to the schema to the SNAPSHOT
|
2021-11-11 08:59:20 +01:00 |
Miriam Baglioni
|
88b73f4f49
|
mergin with branch beta
|
2021-11-10 17:00:52 +01:00 |
Miriam Baglioni
|
c371b23077
|
-
|
2021-11-10 17:00:37 +01:00 |
Alessia Bardi
|
fc8fceaac3
|
create direct link to WT projects as well
|
2021-11-10 14:11:52 +01:00 |
Alessia Bardi
|
6cd91004e3
|
fixed DOI for Wellcome Trust in mapping relationships from Crossref
|
2021-11-09 12:22:57 +01:00 |
Miriam Baglioni
|
9e214ce0eb
|
[BypassAS] addition of OC relations
|
2021-11-09 12:07:19 +01:00 |
Alessia Bardi
|
b9d4f115cc
|
fixed Crossref mappign for SFI projects
|
2021-11-09 12:04:45 +01:00 |
Sandro La Bruzzo
|
6477a40670
|
implement filter of openCitation
|
2021-11-09 11:27:12 +01:00 |
Miriam Baglioni
|
6f7ca539c6
|
[BypassAS] update of results for bipFinder and FOS
|
2021-11-09 11:25:41 +01:00 |
Miriam Baglioni
|
a7d50c499b
|
[BypassAS] prepare FOS subject, test and model for FOS and BipFinder scores
|
2021-11-08 16:44:19 +01:00 |
Antonis Lempesis
|
91354c6068
|
- fetching all context related results
- storing tables as parquet
|
2021-11-08 15:15:46 +02:00 |
Miriam Baglioni
|
94918a673c
|
[Graph DUMP] Fix issue for empty origilaId list
|
2021-11-08 10:25:28 +01:00 |
Claudio Atzori
|
9cb8e4ad21
|
Merge branch 'beta' into hierarchical_orgs_relations
|
2021-11-08 09:40:24 +01:00 |
Miriam Baglioni
|
4c70201412
|
mergin with branch beta
|
2021-11-05 12:29:56 +01:00 |
Miriam Baglioni
|
8442efd8d1
|
[Graph DUMP] Filtering out from the originalIds the id of the result in OpenAIRE
|
2021-11-05 12:29:22 +01:00 |
Claudio Atzori
|
5681e89544
|
Update 'dhp-workflows/dhp-graph-mapper/src/main/resources/eu/dnetlib/dhp/oa/graph/dump/schemas/result_schema.json'
|
2021-11-05 12:18:24 +01:00 |
Miriam Baglioni
|
a22c29fba1
|
[Graph DUMP] Filtering out from the originalIds the id of the result in OpenAIRE
|
2021-11-05 12:08:33 +01:00 |
Miriam Baglioni
|
c10ff6928c
|
[Graph DUMP] add schema of the dump related to the model as in dhp-schemas.2.8.31. Note the measere element at the level of the result has been removed because of issues on where to display it: at the level of the result or at the level of the entity
|
2021-11-05 11:36:21 +01:00 |
Miriam Baglioni
|
0857849a86
|
[Graph DUMP] Remove dump of measure until it will be clear where to put it (at the level of result or at the level of the instance)
|
2021-11-05 11:02:37 +01:00 |
Miriam Baglioni
|
df7ee77c7a
|
[DOIBoost Mapping] removed not needed comments
|
2021-11-04 16:24:07 +01:00 |
Miriam Baglioni
|
de63d29b6f
|
[DOIBoost Mapping] Fix to avoid to produce results with null as identifier (probably due to the filtering function in the factory for the creation of the id)
|
2021-11-04 16:16:40 +01:00 |
Miriam Baglioni
|
d50057b2d9
|
[DOIBoost Mapping] changed the way to create the url for the instance: we use the crooref guidelines https://doi.org/doi
|
2021-11-03 16:59:37 +01:00 |
Miriam Baglioni
|
edf55395e9
|
added test resourse
|
2021-11-03 16:49:30 +01:00 |
Miriam Baglioni
|
d97ea82a29
|
[DOIBoost Mapping] Added test to verify the instance created for Crossref will have just the url related to the doi
|
2021-11-03 16:45:15 +01:00 |
Miriam Baglioni
|
96769b4481
|
[DOIBoost - Mapping] Changed the logic which brought in in the instance urls that should not be there: The urld of the doi in the json is reachable from the root (json/"URL") other urls where added from the links element. Now the mapping from the link element has been removed
|
2021-11-03 16:43:36 +01:00 |
Miriam Baglioni
|
683fe093cf
|
[DOIBoost - Mapping] Remove the addition of the instance to the MAG publication record
|
2021-11-03 15:51:26 +01:00 |
Miriam Baglioni
|
b2bb8d9d79
|
[DOIBoost - Mapping] selecting the url from Crossref containing the doi
|
2021-11-03 15:44:57 +01:00 |
Miriam Baglioni
|
779318961c
|
[DOIBoost - Mapping] removed the url from crossref containing the api.elsevier.com... string in the url
|
2021-11-03 14:38:52 +01:00 |
Miriam Baglioni
|
2480e590d1
|
[DOIBoost - Mapping] changed the type on which to map dissertation from Crossref: from 006 Doctoral thesis to 0044 Thesis since dissertation could be either Doctoral or master thesis
|
2021-11-03 14:25:23 +01:00 |
Miriam Baglioni
|
b9d124bb7c
|
[Enrichment: Propagation through parent-child relationships] Added counters, and changed constraint to verify if filtering out the relation (from classname = harvested to classid != propagation)
|
2021-11-03 13:55:37 +01:00 |