Difference between revisions of "Agenda-2020-06-08"
Jump to navigation
Jump to search
(42 intermediate revisions by 3 users not shown) | |||
Line 3: | Line 3: | ||
Back to https://wiki.egi.eu/wiki/Operations_Meeting | Back to https://wiki.egi.eu/wiki/Operations_Meeting | ||
= General information | = General information = | ||
= Middleware = | = Middleware = | ||
Line 10: | Line 10: | ||
* plans on CentOS8 STARTED | * plans on CentOS8 STARTED | ||
** https://wiki.egi.eu/wiki/Next_middleware_release | ** https://wiki.egi.eu/wiki/Next_middleware_release | ||
== Preview repository == | == Preview repository == | ||
Line 23: | Line 19: | ||
== ARGO/SAM == | == ARGO/SAM == | ||
* | *new metrics added to ARGO_MON_OPERATORS profile on [https://ggus.eu/index.php?mode=ticket_info&ticket_id=147169 May 27th]: eu.egi.CREAMCE-JobSubmit, eu.egi.CREAMCE.WN-Csh, eu.egi.CREAMCE.WN-Softver | ||
** [https://argo-mon.egi.eu/nagios/cgi-bin/status.cgi?servicegroup=SERVICE_CREAM-CE&style=detail results]: 177 endpoints, 15 WARNING (Timeout occurred (900 sec) ), 53 CRITICAL. Success rate 70% (61.6% including the WARNING) | |||
*When eu.egi.CREAMCE.WN-Softver is successful: | |||
** [https://argo-mon | |||
* | |||
CREAM JobOutput OK: retrieved outputSandbox: ['std.err', 'std.out'] | CREAM JobOutput OK: retrieved outputSandbox: ['std.err', 'std.out'] | ||
**** std.err **** | **** std.err **** | ||
**** std.out **** | **** std.out **** | ||
egee01 has UMD 3.14.4 | |||
When it fails: | When it fails: | ||
CREAM JobOutput ERROR [DONE-OK, exitCode=1 ]: retrieved outputSandbox: ['std.err', 'std.out'] | CREAM JobOutput ERROR [DONE-OK, exitCode=1 ]: retrieved outputSandbox: ['std.err', 'std.out'] | ||
**** std.err **** | **** std.err **** | ||
**** std.out **** | **** std.out **** | ||
ERROR: | ERROR: unable to find glite, EMI, LCG or UMD WN version on n1037-amd | ||
* | |||
** | * [https://argo-mon-fedcloud.cro-ngi.hr/nagios/cgi-bin/status.cgi?servicegroup=SERVICE_org.opensciencegrid.htcondorce&style=detail HTCondor-CE probes] included in the ARGO_MON_OPERATORS profile on May 13th: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146949 | ||
** 53 endpoints, 8 CRITICAL, success rate is about 84.9% | |||
== FedCloud == | == FedCloud == | ||
Line 75: | Line 47: | ||
== Feedback from DMSU == | == Feedback from DMSU == | ||
== Verify configuration records == | |||
On a yearly basis, the information registered into GOC-DB need to be verified. | |||
NGIs and RCs have been asked to check them. In particular: | |||
# '''NGI managers should review the people registered and the roles assigned to them, and in particular check the following information:''' | |||
#* E-Mail | |||
#* ROD E-Mail | |||
#* Security E-Mail | |||
:NGI Managers should also review the status of the "not certified" RCs, in according to the [https://wiki.egi.eu/wiki/PROC09#Resource_Center_status_Workflow RC Status Workflow]; | |||
# '''RCs administrators should review the people registered and the roles assigned to them, and in particular check the following information:''' | |||
#* E-Mail | |||
#* telephone numbers | |||
#* CSIRT E-Mail | |||
: RC administrators should also review the information related to the registered service endpoints. | |||
'''The process should be completed by June 22nd.''' | |||
[https://wiki.egi.eu/wiki/Verify_Configuration_Records#2020-05 List of tickets]. | |||
== Monthly Availability/Reliability == | == Monthly Availability/Reliability == | ||
*Under-performed sites in the past A/R reports with issues not yet fixed: | *Under-performed sites in the past A/R reports with issues not yet fixed: | ||
**AfricaArabia: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146877 | |||
***ZA-WITS-CORE: SE hardware problem, machine sent to the vendor; CREAM-CE failures due to a [https://cream-guide.readthedocs.io/en/latest/Releases.html#release-1-16-6 known issue] with the classads library ([https://ggus.eu/index.php?mode=ticket_info&ticket_id=146979 GGUS 146979]) | |||
**AsiaPacific: https://ggus.eu/index.php?mode=ticket_info&ticket_id=142591 | **AsiaPacific: https://ggus.eu/index.php?mode=ticket_info&ticket_id=142591 | ||
***INDIACMS-TIFR: SRM service not published in the BDII ([https://ggus.eu/index.php?mode=ticket_info&ticket_id=142245 142245]), DPM 1.9.0 version ([https://ggus.eu/index.php?mode=ticket_info&ticket_id=145676 145676]); | ***INDIACMS-TIFR: SRM service not published in the BDII ([https://ggus.eu/index.php?mode=ticket_info&ticket_id=142245 142245]), DPM 1.9.0 version ([https://ggus.eu/index.php?mode=ticket_info&ticket_id=145676 145676]); installing a new DPM headnode | ||
**NGI_AEGIS: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146867 | |||
***AEGIS03-ELEF-LEDA: SRM problems, endpoint put out of production until the issues are fixed; CE certificate expired in May. | |||
**NGI_CH: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146870 | |||
***CSCS-LCG2: intermittent igtf failures | |||
***UNIBE-LHEP: issues with DPM upgrade and some network problems; statistics are improving | |||
**NGI_CH: https://ggus.eu/index.php?mode=ticket_info&ticket_id=145818 | **NGI_CH: https://ggus.eu/index.php?mode=ticket_info&ticket_id=145818 | ||
***UNIGE-DPNC: new ARC-CE put in production at the end of April, | ***UNIGE-DPNC: new ARC-CE put in production at the end of April, new failures due to an expired certificate | ||
** | **NGI_DE: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146871 | ||
***GoeGRID: CREAM-CE intermittent failures not affecting ATLAS; failures with ARC-CE | |||
*** | |||
**NGI_IT: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146868 | **NGI_IT: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146868 | ||
***CIRMMP | ***CIRMMP: SRM unreachable due to a firewall, statistics are now improving | ||
***INFN-LECCE | ***INFN-LECCE: CREAMCE and SRM failures, recovering... | ||
**NGI_PL: https://ggus.eu/index.php?mode=ticket_info&ticket_id=140557 | **NGI_PL: https://ggus.eu/index.php?mode=ticket_info&ticket_id=140557 | ||
***TASK: QCG problems was fixed; some CREAM-CE failures; the SRM tests are still failing and the DPM upgrade isn't completed yet | ***TASK: QCG problems was fixed; some CREAM-CE failures; the SRM tests are still failing and the DPM upgrade isn't completed yet | ||
**NGI_UK: | **NGI_UK: | ||
***UKI-NORTHGRID-SHEF-HEP: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146455 needing to re-install the ARC-CE | ***UKI-NORTHGRID-SHEF-HEP: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146455 needing to re-install the ARC-CE | ||
***UKI-SOUTHGRID-SUSX: https://ggus.eu/index.php?mode=ticket_info&ticket_id=144720 Migration from CREAM to ARC, WN migration to CentOS7; SRM to be decommissioned; ARC-CE | ***UKI-SOUTHGRID-SUSX: https://ggus.eu/index.php?mode=ticket_info&ticket_id=144720 Migration from CREAM to ARC, WN migration to CentOS7; SRM to be decommissioned; ARC-CE was failing the IGTF test | ||
**NGI_UK: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146875 | |||
***UKI-LT2-QMUL: intermittent webdav failures, plans to upgrade STORM and LUSTRE in the coming months; suggested to mark the webdav endpoint as "not production" | |||
**ROC_CANADA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146452 | **ROC_CANADA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146452 | ||
***CA-SFU-T2: SLURM problems caused failures to site-BDII freshness check due to some old jobs not properly cancelled; recovered | ***CA-SFU-T2: SLURM problems caused failures to site-BDII freshness check due to some old jobs not properly cancelled; recovered | ||
***CA-WATERLOO-T2: SRM failures not involving production VOs, fixed; some unscheduled downtime affected the the A/R figures; some services still to restore. | ***CA-WATERLOO-T2: SRM failures not involving production VOs, fixed; some unscheduled downtime affected the the A/R figures; some services still to restore. | ||
**NGI_UA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146454 | **NGI_UA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146454 | ||
***UA-KNU: GlueSEUniqueID wasn't published due to missing DNS back-resolving record for IPv6 address of the SE. | ***UA-KNU: GlueSEUniqueID wasn't published due to missing DNS back-resolving record for IPv6 address of the SE. Deployed a new ARC-CE version 6, igtf probe returns UNKNOWN. | ||
*Under-performed sites after 3 consecutive months, under-performed NGIs, QoS violations: (''' | *Under-performed sites after 3 consecutive months, under-performed NGIs, QoS violations: ('''May 2020'''): | ||
** | **NGI_DE: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147313 | ||
*** | ***mainz: some problems in March and April, that could not be fixed easily; in May, the HPC infrastructure was attacked and the whole computer center was shut down. | ||
***wuppertalprod: SRM failures to to a BDII issue, fixed | |||
*** | **NGI_NDGF: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147310 | ||
** | ***FI_HIP_T2: configuration problem that prevents some of the OPS-jobs from running | ||
*** | **NGI_PL: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147311 | ||
***NCBJ-CIS | |||
** | ***PSNC | ||
*** | ***WCSS64 | ||
**NGI_UA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147312 | |||
** | ***UA_ICYB_ARC: cluster storage failure, improving... | ||
*** | |||
**NGI_UA: https://ggus.eu/index.php?mode=ticket_info&ticket_id= | |||
*** | |||
*sites suspended: | *sites suspended: | ||
** | **GRID-UNAM (under-performing, DPM upgrade not completed) | ||
== IPv6 readiness plans == | == IPv6 readiness plans == | ||
Line 134: | Line 119: | ||
== ARC Middleware 5 end of support, migration to ARC 6 == | == ARC Middleware 5 end of support, migration to ARC 6 == | ||
* | * [https://operations-portal.egi.eu/broadcast/archive/2668 EGI Operations Broadcast] | ||
* | * [https://wiki.egi.eu/wiki/PROC16_Decommissioning_of_unsupported_software PROC16 Decommission of unsupported software] | ||
* | * deadline: '''end of July''' | ||
* | * Catalin is in contact with ARC team to get a webinar on ARC administration, scheduled (to be confirmed) for July 6th please contact operations@ for information | ||
* Status | |||
{| class="wikitable sortable" | |||
|- | |||
! Date !! Number of endpoints in BDII !! Number of GGUS tickets !! Issues | |||
|- | |||
| 2020-06-08 || 75 || 42 || Some ARC endpoints publish a timestamp instead of a version like 5.X.Y; we can fairly assume they are ARC6 nightly builds, but we're going to close the corresponding tickets after explicit confirmation from the site admin. | |||
|} | |||
== LCGDM end of support and migration to / enabling of DOME == | == LCGDM end of support and migration to / enabling of DOME == | ||
Line 174: | Line 142: | ||
**VO Managers: to know any barriers in moving away from using SRM protocol https://www.surveymonkey.com/r/2ZMZJVG | **VO Managers: to know any barriers in moving away from using SRM protocol https://www.surveymonkey.com/r/2ZMZJVG | ||
* '''Deployment statistics ( | * '''Deployment statistics (Jun 5th):''' | ||
$ ldapsearch -x -LLL -H ldap://egee-bdii.cnaf.infn.it:2170 -b "GLUE2GroupID=grid,o=glue" '(&(objectClass=GLUE2Manager)(GLUE2ManagerProductName=DPM))' GLUE2ManagerProductVersion GLUE2ManagerID | grep GLUE2ManagerProductVersion | sort | uniq -c | $ ldapsearch -x -LLL -H ldap://egee-bdii.cnaf.infn.it:2170 -b "GLUE2GroupID=grid,o=glue" '(&(objectClass=GLUE2Manager)(GLUE2ManagerProductName=DPM))' GLUE2ManagerProductVersion GLUE2ManagerID | grep GLUE2ManagerProductVersion | sort | uniq -c | ||
1 GLUE2ManagerProductVersion: 1.10.0 | 1 GLUE2ManagerProductVersion: 1.10.0 | ||
65 GLUE2ManagerProductVersion: 1.13.0 | |||
2 GLUE2ManagerProductVersion: 1.13.1 | 2 GLUE2ManagerProductVersion: 1.13.1 | ||
11 GLUE2ManagerProductVersion: 1.13.2 | |||
3 GLUE2ManagerProductVersion: 1.8.10 | 3 GLUE2ManagerProductVersion: 1.8.10 | ||
1 GLUE2ManagerProductVersion: 1.8.9 | 1 GLUE2ManagerProductVersion: 1.8.9 | ||
4 GLUE2ManagerProductVersion: 1.9.0 | 4 GLUE2ManagerProductVersion: 1.9.0 | ||
Line 211: | Line 178: | ||
| https://ggus.eu/index.php?mode=ticket_info&ticket_id=143152 | | https://ggus.eu/index.php?mode=ticket_info&ticket_id=143152 | ||
| SE marked as not production due to some issues that need to be fixed | | SE marked as not production due to some issues that need to be fixed | ||
|- | |- | ||
| INDIACMS-TIFR | | INDIACMS-TIFR | ||
| https://ggus.eu/index.php?mode=ticket_info&ticket_id=142245 | | https://ggus.eu/index.php?mode=ticket_info&ticket_id=142245 | ||
| | | new dpm headnode installed with legacy mode, [https://goc.egi.eu/portal/index.php?Page_Type=Downtime&id=28824 downtime for migration] | ||
|- | |- | ||
| OBSPM | | OBSPM | ||
Line 302: | Line 265: | ||
== Next meeting == | == Next meeting == | ||
July 13th, 2020 https://indico.egi.eu/event/4901/ |
Latest revision as of 14:25, 8 June 2020
Main | EGI.eu operations services | Support | Documentation | Tools | Activities | Performance | Technology | Catch-all Services | Resource Allocation | Security |
Documentation menu: | Home • | Manuals • | Procedures • | Training • | Other • | Contact ► | For: | VO managers • | Administrators |
Back to https://wiki.egi.eu/wiki/Operations_Meeting
General information
Middleware
UMD
- plans on CentOS8 STARTED
Preview repository
- released on 2020-05-08
- Preview 1.27.0 AppDB info (sl6): ARC 6.5.0 and 6.6.0, CVMFS 2.7.2, dCache 5.2.20, frontier-squid 4.11.2, gfal2 2.17.2, xrootd 4.11.3
- Preview 2.27.0 AppDB info (CentOS 7): ARC 6.5.0 and 6.6.0, CVMFS 2.7.2, dCache 5.2.20, frontier-squid 4.11.2, gfal2 2.17.2, xrootd 4.11.3
Operations
ARGO/SAM
- new metrics added to ARGO_MON_OPERATORS profile on May 27th: eu.egi.CREAMCE-JobSubmit, eu.egi.CREAMCE.WN-Csh, eu.egi.CREAMCE.WN-Softver
- results: 177 endpoints, 15 WARNING (Timeout occurred (900 sec) ), 53 CRITICAL. Success rate 70% (61.6% including the WARNING)
- When eu.egi.CREAMCE.WN-Softver is successful:
CREAM JobOutput OK: retrieved outputSandbox: ['std.err', 'std.out'] **** std.err **** **** std.out **** egee01 has UMD 3.14.4
When it fails:
CREAM JobOutput ERROR [DONE-OK, exitCode=1 ]: retrieved outputSandbox: ['std.err', 'std.out'] **** std.err **** **** std.out **** ERROR: unable to find glite, EMI, LCG or UMD WN version on n1037-amd
- HTCondor-CE probes included in the ARGO_MON_OPERATORS profile on May 13th: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146949
- 53 endpoints, 8 CRITICAL, success rate is about 84.9%
FedCloud
Feedback from DMSU
Verify configuration records
On a yearly basis, the information registered into GOC-DB need to be verified. NGIs and RCs have been asked to check them. In particular:
- NGI managers should review the people registered and the roles assigned to them, and in particular check the following information:
- ROD E-Mail
- Security E-Mail
- NGI Managers should also review the status of the "not certified" RCs, in according to the RC Status Workflow;
- RCs administrators should review the people registered and the roles assigned to them, and in particular check the following information:
- telephone numbers
- CSIRT E-Mail
- RC administrators should also review the information related to the registered service endpoints.
The process should be completed by June 22nd.
Monthly Availability/Reliability
- Under-performed sites in the past A/R reports with issues not yet fixed:
- AfricaArabia: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146877
- ZA-WITS-CORE: SE hardware problem, machine sent to the vendor; CREAM-CE failures due to a known issue with the classads library (GGUS 146979)
- AsiaPacific: https://ggus.eu/index.php?mode=ticket_info&ticket_id=142591
- NGI_AEGIS: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146867
- AEGIS03-ELEF-LEDA: SRM problems, endpoint put out of production until the issues are fixed; CE certificate expired in May.
- NGI_CH: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146870
- CSCS-LCG2: intermittent igtf failures
- UNIBE-LHEP: issues with DPM upgrade and some network problems; statistics are improving
- NGI_CH: https://ggus.eu/index.php?mode=ticket_info&ticket_id=145818
- UNIGE-DPNC: new ARC-CE put in production at the end of April, new failures due to an expired certificate
- NGI_DE: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146871
- GoeGRID: CREAM-CE intermittent failures not affecting ATLAS; failures with ARC-CE
- NGI_IT: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146868
- CIRMMP: SRM unreachable due to a firewall, statistics are now improving
- INFN-LECCE: CREAMCE and SRM failures, recovering...
- NGI_PL: https://ggus.eu/index.php?mode=ticket_info&ticket_id=140557
- TASK: QCG problems was fixed; some CREAM-CE failures; the SRM tests are still failing and the DPM upgrade isn't completed yet
- NGI_UK:
- UKI-NORTHGRID-SHEF-HEP: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146455 needing to re-install the ARC-CE
- UKI-SOUTHGRID-SUSX: https://ggus.eu/index.php?mode=ticket_info&ticket_id=144720 Migration from CREAM to ARC, WN migration to CentOS7; SRM to be decommissioned; ARC-CE was failing the IGTF test
- NGI_UK: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146875
- UKI-LT2-QMUL: intermittent webdav failures, plans to upgrade STORM and LUSTRE in the coming months; suggested to mark the webdav endpoint as "not production"
- ROC_CANADA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146452
- CA-SFU-T2: SLURM problems caused failures to site-BDII freshness check due to some old jobs not properly cancelled; recovered
- CA-WATERLOO-T2: SRM failures not involving production VOs, fixed; some unscheduled downtime affected the the A/R figures; some services still to restore.
- NGI_UA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146454
- UA-KNU: GlueSEUniqueID wasn't published due to missing DNS back-resolving record for IPv6 address of the SE. Deployed a new ARC-CE version 6, igtf probe returns UNKNOWN.
- AfricaArabia: https://ggus.eu/index.php?mode=ticket_info&ticket_id=146877
- Under-performed sites after 3 consecutive months, under-performed NGIs, QoS violations: (May 2020):
- NGI_DE: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147313
- mainz: some problems in March and April, that could not be fixed easily; in May, the HPC infrastructure was attacked and the whole computer center was shut down.
- wuppertalprod: SRM failures to to a BDII issue, fixed
- NGI_NDGF: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147310
- FI_HIP_T2: configuration problem that prevents some of the OPS-jobs from running
- NGI_PL: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147311
- NCBJ-CIS
- PSNC
- WCSS64
- NGI_UA: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147312
- UA_ICYB_ARC: cluster storage failure, improving...
- NGI_DE: https://ggus.eu/index.php?mode=ticket_info&ticket_id=147313
- sites suspended:
- GRID-UNAM (under-performing, DPM upgrade not completed)
IPv6 readiness plans
- please provide updates to the IPv6 assessment (ongoing) https://wiki.egi.eu/w/index.php?title=IPV6_Assessment
- if any relevant, information will be summarised at OMB
ARC Middleware 5 end of support, migration to ARC 6
- EGI Operations Broadcast
- PROC16 Decommission of unsupported software
- deadline: end of July
- Catalin is in contact with ARC team to get a webinar on ARC administration, scheduled (to be confirmed) for July 6th please contact operations@ for information
- Status
Date | Number of endpoints in BDII | Number of GGUS tickets | Issues |
---|---|---|---|
2020-06-08 | 75 | 42 | Some ARC endpoints publish a timestamp instead of a version like 5.X.Y; we can fairly assume they are ARC6 nightly builds, but we're going to close the corresponding tickets after explicit confirmation from the site admin. |
LCGDM end of support and migration to / enabling of DOME
- The DPM team has agreed to extend support for security updates for the DPM legacy functionality until 30 September 2019. However, affected service providers should still plan to disable legacy mode well before this date.
- more details: https://wiki.egi.eu/wiki/DPM_End_of_legacy-mode_support
- since DPM 1.10.3 release, it is possible enabling the non-legacy mode DOME (Disk operations Management Engine, see documentation)
- latest DPM/dmlite release 1.13.2 has been also released in UMD 4.10.0
- EGI Operations sent a broadcast to site-admins(first bunch and second bunch) and VO managers with a survey to fill in:
- site-admins: to know the upgrade plans: https://www.surveymonkey.com/r/2PBZSNB
- VO Managers: to know any barriers in moving away from using SRM protocol https://www.surveymonkey.com/r/2ZMZJVG
- Deployment statistics (Jun 5th):
$ ldapsearch -x -LLL -H ldap://egee-bdii.cnaf.infn.it:2170 -b "GLUE2GroupID=grid,o=glue" '(&(objectClass=GLUE2Manager)(GLUE2ManagerProductName=DPM))' GLUE2ManagerProductVersion GLUE2ManagerID | grep GLUE2ManagerProductVersion | sort | uniq -c 1 GLUE2ManagerProductVersion: 1.10.0 65 GLUE2ManagerProductVersion: 1.13.0 2 GLUE2ManagerProductVersion: 1.13.1 11 GLUE2ManagerProductVersion: 1.13.2 3 GLUE2ManagerProductVersion: 1.8.10 1 GLUE2ManagerProductVersion: 1.8.9 4 GLUE2ManagerProductVersion: 1.9.0
Liasing with WLCG to follow-up the upgrade. Opened GGUS tickets asking the following:
- all the sites with older DPM versions than 1.12 are suggested to upgrade to the latest DPM version , following the guide DPM upgrade (chapter 1 Upgrade to DPM 1.10.0 "Legacy Flavour" and chapter 2 Upgrade to DPM 1.10.0 "Dome Flavour")
- DOME and the old LCGDM (srm protocol) will coexist
- Monitoring: sites should enable the monitoring of the HTTP/WebDav and/or GridFTP endpoints
- register the storage service endpoint as webdav and/or globus-GRIDFTP service type, with production flag disabled, providing respectively the URL field and the Extension Properties information as explained in the HOWTO21
- check if the tests are ok
- switch the production flag to "yes"
List of tickets
- DPM upgrade, DOME enabling, and monitoring (33)
- enabling DOME and proper monitoring (6)
- tickets opened by WLCG: 13 and 7
SECMON failures
Several CEs are failing the job submission tests, preventing pakiti to check the vulnerabilities fixes on the WNs.
- original ticket: https://ggus.eu/index.php?mode=ticket_info&ticket_id=143837
- List of tickets to the sites
- https://ggus.eu/index.php?mode=ticket_info&ticket_id=144732
AOB
- Storage accounting: http://goc-accounting.grid-support.ac.uk/storagetest/storagesitesystems.html
- several sites stopped publishing storage accounting records: it needs to investigate on and fix it
Next meeting
July 13th, 2020 https://indico.egi.eu/event/4901/