[logseq-plugin-git:commit] 2025-06-05T19:41:07.930Z

This commit is contained in:
2025-06-05 21:41:08 +02:00
parent 23edf6d6ab
commit cfd25bf303
6304 changed files with 708478 additions and 466 deletions
@@ -0,0 +1,31 @@
# 2021-01-18-1035-Juri-Paper-MSR-with-Wimmer
---
tags:
- '#meetings'
---
- # [[2021-01-18-1035-Juri-Paper-MSR-with-Wimmer]]
**Date**: 2021-01-18
**Time**: 10:35
**Attendend**: #attendance/online
**Project**: #PAPERS/msr11-dataset
**People attended**: #people/juri
_Celano: 🌫 -3°C_
---
Quick chat about the [[paper]] available o Overleaf [MSR2021 - Online LaTeX Editor Overleaf](https://www.overleaf.com/project/5fec90e7f829ed1f031f7af8)
! [[file:///Pasted image 20210118103615.png]]
-
- Creare un dataset di JSON schema.
- {{embed [[HOME]]}}
- 60.000 Json schema
- visto i duplicati per identificarli e rimuoverlo. Ci sono circa 40.000 duplicati
- I Json schema sono conformi a diversi draft di specifica (draft3, draft4, etc.)
- Identificati errori
- Errori Parsing
- Errori di conformance rispetto allo schema
- *E' stato sviluppato un tool per fare questo: wrapper a 2 tool diversi (1 tool copre tre draft)*
- Per quanto riguarda le statistiche
- Estrazioni di unigram per clustering (Juri sta lavorando su questo)
- Luca sta lavorando con MiniHive per calcolare un po di metriche simile al nostro MISE 2014