Skip to content
The Times USA
Menu
  • ABOUT
  • CONTACT
  • LIFESTYLE
  • NATIONAL NEWS
  • BUSINESS
  • INTERNATIONAL NEWS
  • TECHNOLOGY
  • PRICE OF BUSINESS SHOW AUDIOS
Menu

Digging into Data – Unburied Treasure

Posted on June 18, 2020June 19, 2020 by admin

By Elizabeth Thede, Special for The Times USA

Previously, I addressed how a search index like one generated from the program dtSearch® is like a data treasure map, supporting both individual and enterprise-wide concurrent searching with instant hit-highlighted search results. That way, multiple individuals can use the same treasure map at the same time and each arrive at whatever unique X marks the spot that individual is looking for.

But the flip side to this data treasure map is that indexed search can also reveal information that a person who buried it may never have expected would ever come to light. Following are some examples of such buried information that a search engine may uncover. Understanding these examples is an important step to securing your own data.

(1) An end-user might save a file with a filename that doesn’t match the file type. For example, an end-user might mislabel a Microsoft Excel file with a .PDF extension, or mislabel an email archive with a .DOCX extension. However, in parsing a binary format, a search engine like dtSearch will figure out the correct file format specification to apply by looking inside the binary file itself, rather than simply looking at the filename extension. So saving a file with a mismatched filename extension will not have any effect on the ability to uncover text in that file.

(2) An end-user might also nest a file inside another file. An example of that would be an email with a ZIP or RAR attachment and a PDF and Microsoft Word document inside and an Access database fully embedded in the Microsoft Word file. However, a search engine like dtSearch will work its way recursively through such nested attachments, making the inner file contents just as visible as the cover email.

(3) Many “Office”-type applications let you insert obscure metadata where that metadata won’t appear by default when you look at a file in that application. In fact, you may have to click around extensively to find the metadata such that the likelihood of a casual viewer of the file in its native application finding the metadata is very low. However, to a search engine, that metadata is as readily accessible an any other text.

(4) In many applications, an end-user can also hide text by making it the same color as the background underneath the text. As a result, there can be white on white text or black on black text which inside the application itself may not be readily visible. However, to a search engine that text is easily apparent regardless of whether the text color in the application view matches the background color.

(5) An end-user can also slightly misspell a word. Typos in emails are everywhere – at least in my emails. But with fuzzy searching on, dtSearch will automatically sift through those slight typographical errors to find ProJectX when ProJectX is mistyped as ProPectX.

(6) Still another example is credit card numbers that may be buried in files. There is a feature inside of dtSearch that can check if digits presented together may be a valid credit card, even if there is no MasterCard, Visa or American Express insignia, for example.

(7) The final example relates to “image only” PDFs. Have you ever run across a PDF where you try to cut and paste text from it, but you can’t, because it is an image only? A search engine can’t find the text on such a PDF either, because it is an image only. However, a search engine like dtSearch can flag this type of file when it does its indexing, and let you know that you need to run it through an OCR program like Adobe Acrobat. At that point, the text of this file will be buried no more.

dtSearch enterprise and developer products instantly search terabytes of “Office” files, PDFs, emails along with attachments, databases and web-based data. The products can run “on premises” or on online platforms like Azure and AWS. Because dtSearch can instantly search terabytes of data, many customers are large enterprises like Fortune 100 companies and federal, state and international government agencies.

However, in addition to enterprise-level search, dtSearch also lets you search your own documents, emails and the like. Please go to dtSearch.com, and download a fully-functional 30-day evaluation version to instantly search terabytes of your own data.

RELATED: Kevin Price of the Price of Business show discusses the topic with Thede on a recent interview.

You Might Also Like...

  • Data Privacy's Importance to Americans

    By the Price of Business Show, Hosted by Kevin Price.  The Price of Business is a media…

  • Beyond Boolean Search

    By Elizabeth Thede, Special for The Time USA   Many people have heard of Boolean…

  • Bad Data Practice Hurts the Bottom Line

    A new report from Dun & Bradstreet reveals businesses are missing revenue opportunities and losing customers due…

  • Data Management and Implementation Services for RingLead Customers

    DemandGen International, Inc., a world-class team of digital transformation and technology experts, today announced a…

  • Talent Trends Data Shows Strong Global Outlook for the New Year

    C-suite and human capital leaders surveyed around the world continue to feel positive about their…

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

VIDEO: This Week’s Best of our Network

https://www.youtube.com/watch?v=cpTRZlvuuuI

GDPR Compliance

USABR does not collect data on its visitors.  For more information visit: https://www.usabusinessradio.com/contact-us/

Contact

Contact articles@usabusinessradio.net for more information on articles on this site. BMuyco@usabusinessradio.net for all other information.

Recent Articles

  • Working Together To Ensure No Kid Fights Cancer Alone
  • Urgent Relief, Lasting Impact
  • The Life Expectancy Rates of Truck Drivers Will Shock You
  • The Secret Core of a Flawless Manicure: Kodi Rubber Base
  • US Stablecoin Regulation: A New Era Under the Genius Act

Also in TTUSA

  • Featuring the World’s Largest Emerald Specimen Collection
  • Making Video Presentations from ABD-Video
  • How the First YouTuber to Talk about Drop Shipping, Dan Dasilva, Is Building a Real Lasting Impact in the Tech World
  • Improve MacBook Experience with these Hidden Settings
  • How Can Orthodontics Assist in Dental Health?

RSS The Daily Blaze

  • Why So Few People Understand the Value a Business Broker Brings to Deals
  • What’s Going On With Big Cities and Trash?
  • Trump’s White House East Wing Project Is Unprecedented on Every Front
  • Moving to San Mateo: What To Consider When Transporting Furniture in a Coastal Climate
  • Security Expert on Latest Schemes in Business Email Compromises

RSS USA Business Radio

  • Leading Business Broker Answers the Question, “When Should I Prepare To Sell?”
  • Handling Succession, Estate Planning & Litigation
  • Leading Media Authority on the Selling of CNN
  • Animal Rescue That Gives Pets a Meaningful Life, Not Just Survival
  • Reserve Your Seat to Space With Virgin Galactic

RSS USA Daily Times

  • The Evolution of Ultra-Luxe Travel and the Vision of Onirikos
  • 2026 Luxury Travel: Hyper-Personalized, Cooler, Screen-Inspired, and World-Class Sports Experiences
  • New Winners Circles for Retired Thoroughbreds Thru Thoroughbred Rescue
  • Healthy Alternatives to Your Favorite Candy Bars
  • Marrakech’s Majestic Stays: Four Icons of Luxury

RSS USA Daily Chronicles.

  • The Power of Calorie Density: Why What You Eat Matters As Much as How Much
  • Saving Kittens and Cats Through Adoption
  • If We Can Save Butterflies, We Can Save Ourselves
  • Don’t Rely on Third-Party Weight Loss Programs
  • Dr. Michael Jacobson on the Hidden Dangers of Ultra-Processed Foods

RSS Price of Business

RSS US Daily Review

  • When Failure Teaches Leadership With Dr. Don McNeeley
  • Shutdown Persists Until Healthcare Awareness Hits
  • When Crimes Seem Built for Netflix
  • The Language of Job Loss: From Orwell to HR-Speak
  • The Important Work of Wildlife Rehabilitation

PoB Digital Network

US Daily Review

USA Business Radio

USA Daily Chronicles

USA Daily Times

The Daily Blaze

The Times USA

Price of Business

© 2025 The Times USA | Powered by Superbs Personal Blog theme