[HN Gopher] AI could help make Wikipedia entries more accurate
       ___________________________________________________________________
        
       AI could help make Wikipedia entries more accurate
        
       Author : panabee
       Score  : 10 points
       Date   : 2022-07-11 15:58 UTC (7 hours ago)
        
 (HTM) web link (tech.fb.com)
 (TXT) w3m dump (tech.fb.com)
        
       | juve1996 wrote:
       | "could"
       | 
       | Tell me when you "can."
        
         | camdat wrote:
         | Did you bother to read the article?
         | 
         | >Building on Meta AI's research and advancements, we've
         | developed the first model _capable of_ automatically scanning
         | hundreds of thousands of citations at once to check whether
         | they truly support the corresponding claims.
        
           | adhesive_wombat wrote:
           | > check whether they truly support
           | 
           | Check whether they statistically _might_ support, more like.
           | 
           | No AI can tell you if an article does or doesn't support
           | something with complete confidence.
        
           | [deleted]
        
       | permo-w wrote:
       | I object to this
        
       | jka wrote:
       | Trying to ignore any hype, lofty sci-fi ideas, or potential
       | philosophical questions for a moment: roughly speaking, it sounds
       | like this is a search engine, for use in a neat and thought-
       | provoking use case.
       | 
       | There's an architecture diagram[1] alongside the source code, and
       | my summary would be:
       | 
       | - The system has in-house web indexes built from Common Crawl[2]
       | data
       | 
       | - The system receives snippets of text from Wikipedia and
       | determines whether existing citations exist and whether they are
       | valid
       | 
       | - If no valid citation exists, then the system performs queries
       | against the indexes to find relevant URLs
       | 
       | It'd be interesting to learn how this approach fares compared to
       | pasting the relevant paragraphs of text into search engines and
       | excluding site:wikipedia.org from the results.
       | 
       | Something about feedback loops and data quality makes me wary
       | that too much application of automated systems like this would
       | lead to a degradation of content quality (each updated copy an
       | imperfect translation or reference to an existing one).
       | 
       | [1] -
       | https://github.com/facebookresearch/side/tree/a595fb09c85233...
       | 
       | [2] - https://commoncrawl.org/
        
       | sdfhdhjdw3 wrote:
       | FB is a cancer, lets keep it far away from wikipedia.
        
       | macrolocal wrote:
       | Nb. http://xowa.org/ is one way to archive Wikipedia.
        
       ___________________________________________________________________
       (page generated 2022-07-11 23:02 UTC)