Miriam Baglioni
|
20de75ca64
|
[Measures] removed typo
|
2022-04-21 12:14:03 +02:00 |
Miriam Baglioni
|
bebb2a0560
|
Merge branch 'eosc_dimitris' of https://code-repo.d4science.org/D-Net/dnet-hadoop into eosc_dimitris
|
2022-04-21 12:10:19 +02:00 |
Miriam Baglioni
|
b61efd613b
|
[Measures] addressed comments in the PR
|
2022-04-21 12:09:37 +02:00 |
Miriam Baglioni
|
d012d125d7
|
[EOSCTag] -
|
2022-04-21 12:02:09 +02:00 |
Claudio Atzori
|
88acad76f9
|
Merge branch 'beta' into eosc_dimitris
|
2022-04-21 12:00:03 +02:00 |
Miriam Baglioni
|
c304657d91
|
[Measures] put the logic in common, no need to change the schema
|
2022-04-21 11:27:26 +02:00 |
Miriam Baglioni
|
5295effc96
|
[Measures] fixed issue
|
2022-04-20 16:20:40 +02:00 |
Miriam Baglioni
|
a38f0f5ea7
|
mergin with branch beta
|
2022-04-20 15:44:18 +02:00 |
Miriam Baglioni
|
dbfbe8841a
|
[Clean Context] changed the description in input parameters
|
2022-04-20 15:41:03 +02:00 |
Miriam Baglioni
|
5feae77937
|
[Measures] last changes to accomodate tests
|
2022-04-20 15:13:09 +02:00 |
Miriam Baglioni
|
869407c6e2
|
[Measures] added new measure (usagecounts) as action set. Measure added at the level of the result. Ref #7587
|
2022-04-20 14:02:05 +02:00 |
Michele Artini
|
c96a8613f8
|
update SQL queries
|
2022-04-20 12:07:49 +02:00 |
Michele Artini
|
4314db55c8
|
migration to services: update sql queries
|
2022-04-19 15:05:02 +02:00 |
Sandro La Bruzzo
|
d5b29d96a7
|
fix merging in crossrefAggregator which creates dataInfo null
|
2022-04-14 11:07:04 +02:00 |
Claudio Atzori
|
b93a141d6c
|
[Doiboost] fixed fundingReference extraction from the Crossref records
|
2022-04-12 10:26:05 +02:00 |
Claudio Atzori
|
73c172926a
|
[Doiboost] fixed fundingReference extraction from the Crossref records
|
2022-04-12 10:25:42 +02:00 |
Claudio Atzori
|
48b580b45c
|
[graph enrichment] fixed country_propagation oozie workflow definition, parameter saveGraph is not needed anymore by the SparkCountryPropagationJob
|
2022-04-11 08:52:36 +02:00 |
Claudio Atzori
|
21f32b83c6
|
[graph enrichment] fixed country_propagation oozie workflow definition, parameter saveGraph is not needed anymore by the SparkCountryPropagationJob
|
2022-04-11 08:52:12 +02:00 |
Claudio Atzori
|
4eff7856f5
|
Merge pull request '[stats-wf] computing stats in each step' (#210) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#210
|
2022-04-08 14:21:01 +02:00 |
Claudio Atzori
|
c26222623f
|
[maven-release-plugin] prepare for next development iteration
|
2022-04-07 13:32:22 +02:00 |
Claudio Atzori
|
86585a6b27
|
[maven-release-plugin] prepare release dhp-1.2.4
|
2022-04-07 13:32:19 +02:00 |
Claudio Atzori
|
ad85d88eaf
|
[maven-release-plugin] rollback the release of dhp-1.2.4
|
2022-04-07 13:28:35 +02:00 |
Claudio Atzori
|
598e11dfd7
|
[maven-release-plugin] prepare for next development iteration
|
2022-04-07 13:27:02 +02:00 |
Claudio Atzori
|
db3d9877a5
|
[maven-release-plugin] prepare release dhp-1.2.4
|
2022-04-07 13:26:58 +02:00 |
Enrico Ottonello
|
9a0ca0296a
|
added mobidb constants configuration and test
|
2022-04-07 13:17:50 +02:00 |
Claudio Atzori
|
3bba6d6e38
|
[maven-release-plugin] rollback the release of dhp-1.2.4
|
2022-04-07 12:23:17 +02:00 |
Claudio Atzori
|
2ac2d928bd
|
[maven-release-plugin] prepare for next development iteration
|
2022-04-07 12:18:47 +02:00 |
Claudio Atzori
|
85bc722ff4
|
[maven-release-plugin] prepare release dhp-1.2.4
|
2022-04-07 12:18:43 +02:00 |
Claudio Atzori
|
bc05b6168a
|
[maven-release-plugin] rollback the release of dhp-1.2.4
|
2022-04-07 11:49:06 +02:00 |
Claudio Atzori
|
505420fd61
|
[maven-release-plugin] prepare for next development iteration
|
2022-04-07 11:34:06 +02:00 |
Claudio Atzori
|
66e718981e
|
[maven-release-plugin] prepare release dhp-1.2.4
|
2022-04-07 11:34:02 +02:00 |
Claudio Atzori
|
05fafa1408
|
[graph raw] avoid NPEs importing datasource consent fields
|
2022-04-06 15:23:50 +02:00 |
Enrico Ottonello
|
d0df02062c
|
raised runtimeexception on record without title or url
|
2022-04-06 13:19:58 +02:00 |
Enrico Ottonello
|
7fc5b97871
|
skipped record without title or url
|
2022-04-06 13:08:48 +02:00 |
Enrico Ottonello
|
a203c33693
|
added disprot constants configuration
|
2022-04-06 12:48:24 +02:00 |
Enrico Ottonello
|
afb46d71f7
|
added subjects to disprot output
|
2022-04-06 12:22:10 +02:00 |
Antonis Lempesis
|
c442c91f89
|
computing stats in each step
|
2022-04-06 12:40:02 +03:00 |
Claudio Atzori
|
8c457f1b2c
|
conflicts resolved, merged from beta
|
2022-04-06 10:27:52 +02:00 |
Miriam Baglioni
|
e77d104951
|
[OC] added / to workflow path
|
2022-04-05 15:07:11 +02:00 |
Enrico Ottonello
|
98178b3165
|
custom deserializer for property value type working for both ped and disprot
|
2022-04-05 10:47:17 +02:00 |
Miriam Baglioni
|
79336d46c5
|
[Clean Context] first naive implementation of a functionality to clean not wanted contextes from one result. This implementation simply verifies the main title of the results start with a given string
|
2022-04-04 15:52:31 +02:00 |
Claudio Atzori
|
873369af1c
|
Merge pull request '[stats wf] added apcs in monitor db' (#207) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#207
|
2022-03-29 15:40:20 +02:00 |
Antonis Lempesis
|
7112806a73
|
views cannot be stored as parquet...
|
2022-03-29 16:37:29 +03:00 |
Antonis Lempesis
|
fff0b3cc19
|
added apcs in monitor db
|
2022-03-29 14:15:31 +03:00 |
Claudio Atzori
|
de85367695
|
Merge pull request '[stats wf] fix: views cannot be stored as parquet...' (#206) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#206
|
2022-03-29 12:51:02 +02:00 |
Antonis Lempesis
|
ee24f3eb2c
|
views cannot be stored as parquet...
|
2022-03-29 13:47:48 +03:00 |
Sandro La Bruzzo
|
1b11010169
|
minor fix
|
2022-03-29 10:59:14 +02:00 |
Claudio Atzori
|
0a0ae84c22
|
[graph raw] DOI based instance URLs on https
|
2022-03-29 10:52:58 +02:00 |
Claudio Atzori
|
9fa3dd78fe
|
Merge pull request '[stats wf] various fixes, organization ids for inst. dashboard' (#205) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#205
|
2022-03-28 22:03:49 +02:00 |
Claudio Atzori
|
96aa2a5d0d
|
Merge branch 'beta' into instance_group_by_url
|
2022-03-28 09:23:52 +02:00 |
Claudio Atzori
|
741bc99c47
|
Merge branch 'beta' into datasource_pdf_consent
|
2022-03-28 09:20:48 +02:00 |
Claudio Atzori
|
61319b2e83
|
updated dhp-schema version; set entity-level dataInfo before & after merging the fields from the group of duplicates
|
2022-03-25 16:38:33 +01:00 |
Antonis Lempesis
|
d8503cd191
|
added moooar organizations
|
2022-03-24 14:02:36 +02:00 |
Miriam Baglioni
|
7b8f85692e
|
[Enrichment country] fixed issues with parameters and workflow args
|
2022-03-23 17:20:23 +01:00 |
Claudio Atzori
|
48d32466e4
|
instances grouped by URL expose only one refereed
|
2022-03-23 14:52:03 +01:00 |
Claudio Atzori
|
f10066547b
|
increased spark.sql.shuffle.partitions in affiliation_from_semrel_propagation
|
2022-03-23 12:22:26 +01:00 |
Claudio Atzori
|
43733c1a18
|
Merge branch 'beta' of https://code-repo.d4science.org/D-Net/dnet-hadoop into beta
|
2022-03-23 12:14:27 +01:00 |
Enrico Ottonello
|
f11dfc51f7
|
fix resolved url format, added alternate identifier from original pid
|
2022-03-22 16:39:21 +01:00 |
Antonis Lempesis
|
62f91b0869
|
cleanup
|
2022-03-22 16:17:49 +02:00 |
Antonis Lempesis
|
2e8394ecf8
|
creating aaall tables as parquet
|
2022-03-22 16:16:08 +02:00 |
Antonis Lempesis
|
dcfbeb8142
|
yet more typos
|
2022-03-21 12:36:03 +02:00 |
Miriam Baglioni
|
89fd275480
|
[HostedByMap] added left over from PR and fixed issue on workflow
|
2022-03-21 09:54:45 +01:00 |
Enrico Ottonello
|
afe84c4244
|
added subjects to oaf generation
|
2022-03-18 18:10:39 +01:00 |
Enrico Ottonello
|
db831e6f43
|
removed dynamic allocation on wf
|
2022-03-18 17:43:53 +01:00 |
Enrico Ottonello
|
861f2a3306
|
added titles merging title page and protein identifier
|
2022-03-18 14:51:57 +01:00 |
Enrico Ottonello
|
f43bfdb594
|
added subjects
|
2022-03-17 19:24:07 +01:00 |
miconis
|
c763aded70
|
dependency updated to the new pace-core version
|
2022-03-16 16:41:50 +01:00 |
Enrico Ottonello
|
3ef5eec3a6
|
added bmuse and rdfconverter modules - added repository for bmuse jars
|
2022-03-16 12:07:36 +01:00 |
Enrico Ottonello
|
41284ec2f9
|
retrieving vocabulary terms from nquads
|
2022-03-16 11:26:50 +01:00 |
Enrico Ottonello
|
e53a606afc
|
added date of collection, resource type as workflow parameter
|
2022-03-15 17:36:48 +01:00 |
miconis
|
c959639bd5
|
dependency updated to the new pace-core version
|
2022-03-15 16:33:03 +01:00 |
Miriam Baglioni
|
0f7d8ca2e0
|
[HostedByMap] change on master to align to PR 201 on beta merged as 9f3036c847
|
2022-03-11 15:16:02 +01:00 |
Claudio Atzori
|
f430029596
|
cleanup
|
2022-03-11 14:28:28 +01:00 |
Miriam Baglioni
|
12de9acb0d
|
[Country Propagation] left out from previous commit
|
2022-03-11 14:17:02 +01:00 |
Miriam Baglioni
|
2fbb35ade5
|
mergin with branch beta
|
2022-03-11 13:58:10 +01:00 |
Miriam Baglioni
|
4437f9345d
|
[Country Propagation] left out from previous commit
|
2022-03-11 13:57:47 +01:00 |
Miriam Baglioni
|
2b643059fa
|
[Country Propagation] changed the logic to get the collectedfrom at the result level. To fix issue when no instance is created for a result that should have the country associated. Change the code to use spark instead of hive to prepare the data needed for the propagation step. Added new tests for the intermediate steps and new verification for the propagation itself
|
2022-03-11 13:56:48 +01:00 |
Claudio Atzori
|
f25407bbe2
|
added mapping for datasource consent fields to integrate them in the graph
|
2022-03-11 09:32:42 +01:00 |
Miriam Baglioni
|
2c5087d55a
|
[HostedByMap] download of doaj from json, modification of test resources, deletion of class no more needed for the CSV download
|
2022-03-04 15:18:21 +01:00 |
Miriam Baglioni
|
5d608d6291
|
[HostedByMap] changed the model to include also oaStart date and review process that could be possibly used in the future
|
2022-03-04 11:06:09 +01:00 |
Miriam Baglioni
|
b7c2340952
|
[HostedByMap - DOIBoost] changed to use code moved to common since used also from hostedbymap now
|
2022-03-04 11:05:23 +01:00 |
Miriam Baglioni
|
8a41f63348
|
[HostedByMap] update to download the json instead of the csv
|
2022-03-04 10:38:43 +01:00 |
Miriam Baglioni
|
44b0c03080
|
[HostedByMap] update to download the json instead of the csv
|
2022-03-04 10:37:59 +01:00 |
Enrico Ottonello
|
bd37f14941
|
added working ocean configuration
|
2022-03-03 14:38:21 +01:00 |
Enrico Ottonello
|
29ee1b9d82
|
added datasource key to workflow parameter to properly choose collected from and id values
|
2022-03-03 12:31:29 +01:00 |
Antonis Lempesis
|
ad78e505da
|
yet another fix
|
2022-03-03 12:28:12 +02:00 |
Enrico Ottonello
|
e57216a1fa
|
added oozie workflow to generate bioschema dataset on hdfs
|
2022-03-02 16:58:10 +01:00 |
Miriam Baglioni
|
3be8737c32
|
[graph-stats] fixed query after the change in the indicator table related to PR#200
|
2022-03-02 14:09:05 +01:00 |
Miriam Baglioni
|
3970651ee1
|
Merge pull request 'fixed query after the change in the indicator table' (#200) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#200
|
2022-03-02 14:05:58 +01:00 |
Antonis Lempesis
|
efeeebfee1
|
fixed query after the change in the indicator table
|
2022-03-02 13:29:25 +02:00 |
Enrico Ottonello
|
f28d7e3b9d
|
added spark dataset creation
|
2022-03-02 12:12:37 +01:00 |
Enrico Ottonello
|
8f281846a4
|
update test file
|
2022-02-28 13:37:53 +01:00 |
Enrico Ottonello
|
f833de8a75
|
added relatedIdentifierType, fix IsCitedBy value
|
2022-02-28 13:37:28 +01:00 |
Enrico Ottonello
|
7f9636ef00
|
added alternateIdentifiers to oaf
|
2022-02-25 14:42:08 +01:00 |
Claudio Atzori
|
580d904aae
|
manually merging PR#199 D-Net/dnet-hadoop#199
|
2022-02-25 12:22:50 +01:00 |
Claudio Atzori
|
1932a65d1c
|
Merge pull request '[Stats wf] sprint 6 indicators' (#198) from antonis.lempesis/dnet-hadoop:beta into beta
Reviewed-on: D-Net/dnet-hadoop#198
|
2022-02-25 12:09:18 +01:00 |
Miriam Baglioni
|
f5b0a6f89c
|
[master to beta] fixed issues in test files
|
2022-02-25 10:21:57 +01:00 |
miconis
|
8991d097b4
|
bug fix in the DedupRecordFactory, DataInfo set before merge
|
2022-02-24 17:13:12 +01:00 |
miconis
|
fe1c966cbf
|
Merge branch 'master_202203' of code-repo.d4science.org:D-Net/dnet-hadoop into master_202203
|
2022-02-24 17:08:38 +01:00 |
miconis
|
b0f369dc78
|
bug fix in the DedupRecordFactory, DataInfo set before merge
|
2022-02-24 17:08:24 +01:00 |