YouTube2Text

Deduplicating & Merging Your Library | HubMeta Tutorial #6 — Transcript

by HubMeta · 443 words · 62 segments · language en · Watch on YouTube

Full transcript

  1. 0:14All right. Now that our library has the
  2. 0:18articles that we want, we want to remove
  3. 0:21duplicates from that. We can go click on
  4. 0:24ddup and add to project from our import
  5. 0:28page or go to pull and streamline page.
  6. 0:32Uh the first time you come on here it
  7. 0:34might take some time to load all of your
  8. 0:37files. If you saw that you have files in
  9. 0:39the import page they are not showing up
  10. 0:42here. Do a refresh of the page and uh
  11. 0:46you know all your files will appear.
  12. 0:48What we need to do is actually very
  13. 0:50simple. We will select the files that we
  14. 0:53want to remove duplicates from. So you
  15. 0:55see I have selected all of the my tree
  16. 0:58files and I will click on dduplicate
  17. 1:02tree files. This will run a process in
  18. 1:05the background
  19. 1:07to remove the duplicates. If you wanted
  20. 1:09to report how this dduplication works,
  21. 1:13so you can click on this. We have very
  22. 1:15clearly explained how our dduplication
  23. 1:18algorithm works. If you had any
  24. 1:21questions about this, we are more than
  25. 1:23happy to share more details about that,
  26. 1:26maybe now would be a good time to
  27. 1:27explain that all the decisions that we
  28. 1:30have made in building hub meta. We have
  29. 1:33tried our best to make it transparent.
  30. 1:35If anything that we do in the background
  31. 1:38is not very clear, always feel free to
  32. 1:42reach out to myself, send an email to
  33. 1:45supportshubmeda.com
  34. 1:46and ask your question. Our goal is to
  35. 1:49create an open science platform and to
  36. 1:52provide everything transparently so you
  37. 1:55can report it and create a reproducible
  38. 1:58systematic review. Now my dduplication
  39. 2:01process is finished. You can see that
  40. 2:05out of the 29,000 records around 30,000
  41. 2:09it found
  42. 2:117700 duplicates the final count is final
  43. 2:16records and it did that in 1 minutes.
  44. 2:19Here is an important piece. What we have
  45. 2:22done after this is that we have created
  46. 2:24a dduplicated version down here. It is
  47. 2:27not yet uploaded into your project. your
  48. 2:31project count is still zero because we
  49. 2:35want you to deliberately select the
  50. 2:37files that you want to go into your
  51. 2:40project. So I will select this
  52. 2:41dduplicated file not the original files
  53. 2:45anymore and then click on merge to
  54. 2:48project. That is how we start
  55. 2:51everything. So during this merge
  56. 2:54process, it will read all the RAIS. It
  57. 2:56will see if it can uh find any issues in
  58. 3:00any of them, you know, if it can
  59. 3:02supplement it with more information and
  60. 3:06then uh it will all be added to your
  61. 3:09project. See you in the next video to
  62. 3:12talk about more steps in the pool stage.

About this transcript

This page contains the full transcript of Deduplicating & Merging Your Library | HubMeta Tutorial #6 by HubMeta, generated from the public captions YouTube serves with the video. The transcript has 443 words across 62 segments, with the original timestamps preserved so you can click any line to jump to that moment in the embedded player.

What you can do with it

Use the transcript to take notes, quote the speaker, build a study guide, generate a summary with ChatGPT or Claude via the YouTube Summary tool, or export it as a timed subtitle file with YouTube to SRT. You can also re-open it in the transcriber to translate the transcript into 100+ languages.

Free YouTube transcript tool

YouTube2Text is a free YouTube transcript generator — no signup, no daily limit. Paste any YouTube link and get the full transcript instantly, with timestamps, click-to-jump, translation to 100+ languages, AI prompts for ChatGPT, Claude, and Gemini, and exports to TXT, SRT, VTT, or Markdown.