Deduplicating & Merging Your Library | HubMeta Tutorial #6 — Transcript
Full transcript
- 0:14All right. Now that our library has the
- 0:18articles that we want, we want to remove
- 0:21duplicates from that. We can go click on
- 0:24ddup and add to project from our import
- 0:28page or go to pull and streamline page.
- 0:32Uh the first time you come on here it
- 0:34might take some time to load all of your
- 0:37files. If you saw that you have files in
- 0:39the import page they are not showing up
- 0:42here. Do a refresh of the page and uh
- 0:46you know all your files will appear.
- 0:48What we need to do is actually very
- 0:50simple. We will select the files that we
- 0:53want to remove duplicates from. So you
- 0:55see I have selected all of the my tree
- 0:58files and I will click on dduplicate
- 1:02tree files. This will run a process in
- 1:05the background
- 1:07to remove the duplicates. If you wanted
- 1:09to report how this dduplication works,
- 1:13so you can click on this. We have very
- 1:15clearly explained how our dduplication
- 1:18algorithm works. If you had any
- 1:21questions about this, we are more than
- 1:23happy to share more details about that,
- 1:26maybe now would be a good time to
- 1:27explain that all the decisions that we
- 1:30have made in building hub meta. We have
- 1:33tried our best to make it transparent.
- 1:35If anything that we do in the background
- 1:38is not very clear, always feel free to
- 1:42reach out to myself, send an email to
- 1:45supportshubmeda.com
- 1:46and ask your question. Our goal is to
- 1:49create an open science platform and to
- 1:52provide everything transparently so you
- 1:55can report it and create a reproducible
- 1:58systematic review. Now my dduplication
- 2:01process is finished. You can see that
- 2:05out of the 29,000 records around 30,000
- 2:09it found
- 2:117700 duplicates the final count is final
- 2:16records and it did that in 1 minutes.
- 2:19Here is an important piece. What we have
- 2:22done after this is that we have created
- 2:24a dduplicated version down here. It is
- 2:27not yet uploaded into your project. your
- 2:31project count is still zero because we
- 2:35want you to deliberately select the
- 2:37files that you want to go into your
- 2:40project. So I will select this
- 2:41dduplicated file not the original files
- 2:45anymore and then click on merge to
- 2:48project. That is how we start
- 2:51everything. So during this merge
- 2:54process, it will read all the RAIS. It
- 2:56will see if it can uh find any issues in
- 3:00any of them, you know, if it can
- 3:02supplement it with more information and
- 3:06then uh it will all be added to your
- 3:09project. See you in the next video to
- 3:12talk about more steps in the pool stage.
About this transcript
This page contains the full transcript of Deduplicating & Merging Your Library | HubMeta Tutorial #6 by HubMeta, generated from the public captions YouTube serves with the video. The transcript has 443 words across 62 segments, with the original timestamps preserved so you can click any line to jump to that moment in the embedded player.
What you can do with it
Use the transcript to take notes, quote the speaker, build a study guide, generate a summary with ChatGPT or Claude via the YouTube Summary tool, or export it as a timed subtitle file with YouTube to SRT. You can also re-open it in the transcriber to translate the transcript into 100+ languages.
Free YouTube transcript tool
YouTube2Text is a free YouTube transcript generator — no signup, no daily limit. Paste any YouTube link and get the full transcript instantly, with timestamps, click-to-jump, translation to 100+ languages, AI prompts for ChatGPT, Claude, and Gemini, and exports to TXT, SRT, VTT, or Markdown.