Skip to content
The Times USA
Menu
  • ABOUT
  • CONTACT
  • LIFESTYLE
  • NATIONAL NEWS
  • BUSINESS
  • INTERNATIONAL NEWS
  • TECHNOLOGY
  • PRICE OF BUSINESS SHOW AUDIOS
Menu

Digging into Data – Unburied Treasure

Posted on June 18, 2020June 19, 2020 by admin

By Elizabeth Thede, Special for The Times USA

Previously, I addressed how a search index like one generated from the program dtSearch® is like a data treasure map, supporting both individual and enterprise-wide concurrent searching with instant hit-highlighted search results. That way, multiple individuals can use the same treasure map at the same time and each arrive at whatever unique X marks the spot that individual is looking for.

But the flip side to this data treasure map is that indexed search can also reveal information that a person who buried it may never have expected would ever come to light. Following are some examples of such buried information that a search engine may uncover. Understanding these examples is an important step to securing your own data.

(1) An end-user might save a file with a filename that doesn’t match the file type. For example, an end-user might mislabel a Microsoft Excel file with a .PDF extension, or mislabel an email archive with a .DOCX extension. However, in parsing a binary format, a search engine like dtSearch will figure out the correct file format specification to apply by looking inside the binary file itself, rather than simply looking at the filename extension. So saving a file with a mismatched filename extension will not have any effect on the ability to uncover text in that file.

(2) An end-user might also nest a file inside another file. An example of that would be an email with a ZIP or RAR attachment and a PDF and Microsoft Word document inside and an Access database fully embedded in the Microsoft Word file. However, a search engine like dtSearch will work its way recursively through such nested attachments, making the inner file contents just as visible as the cover email.

(3) Many “Office”-type applications let you insert obscure metadata where that metadata won’t appear by default when you look at a file in that application. In fact, you may have to click around extensively to find the metadata such that the likelihood of a casual viewer of the file in its native application finding the metadata is very low. However, to a search engine, that metadata is as readily accessible an any other text.

(4) In many applications, an end-user can also hide text by making it the same color as the background underneath the text. As a result, there can be white on white text or black on black text which inside the application itself may not be readily visible. However, to a search engine that text is easily apparent regardless of whether the text color in the application view matches the background color.

(5) An end-user can also slightly misspell a word. Typos in emails are everywhere – at least in my emails. But with fuzzy searching on, dtSearch will automatically sift through those slight typographical errors to find ProJectX when ProJectX is mistyped as ProPectX.

(6) Still another example is credit card numbers that may be buried in files. There is a feature inside of dtSearch that can check if digits presented together may be a valid credit card, even if there is no MasterCard, Visa or American Express insignia, for example.

(7) The final example relates to “image only” PDFs. Have you ever run across a PDF where you try to cut and paste text from it, but you can’t, because it is an image only? A search engine can’t find the text on such a PDF either, because it is an image only. However, a search engine like dtSearch can flag this type of file when it does its indexing, and let you know that you need to run it through an OCR program like Adobe Acrobat. At that point, the text of this file will be buried no more.

dtSearch enterprise and developer products instantly search terabytes of “Office” files, PDFs, emails along with attachments, databases and web-based data. The products can run “on premises” or on online platforms like Azure and AWS. Because dtSearch can instantly search terabytes of data, many customers are large enterprises like Fortune 100 companies and federal, state and international government agencies.

However, in addition to enterprise-level search, dtSearch also lets you search your own documents, emails and the like. Please go to dtSearch.com, and download a fully-functional 30-day evaluation version to instantly search terabytes of your own data.

RELATED: Kevin Price of the Price of Business show discusses the topic with Thede on a recent interview.

You Might Also Like...

  • Data Privacy's Importance to Americans

    By the Price of Business Show, Hosted by Kevin Price.  The Price of Business is a media…

  • Beyond Boolean Search

    By Elizabeth Thede, Special for The Time USA   Many people have heard of Boolean…

  • Bad Data Practice Hurts the Bottom Line

    A new report from Dun & Bradstreet reveals businesses are missing revenue opportunities and losing customers due…

  • Data Management and Implementation Services for RingLead Customers

    DemandGen International, Inc., a world-class team of digital transformation and technology experts, today announced a…

  • Talent Trends Data Shows Strong Global Outlook for the New Year

    C-suite and human capital leaders surveyed around the world continue to feel positive about their…

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

VIDEO: This Week’s Best of our Network

https://www.youtube.com/watch?v=GMS9ouZIDqw

GDPR Compliance

USABR does not collect data on its visitors.  For more information visit: https://www.usabusinessradio.com/contact-us/

Contact

Contact articles@usabusinessradio.net for more information on articles on this site. BMuyco@usabusinessradio.net for all other information.

Recent Articles

  • Rep. Haley Stevens Makes the Case for Accountability at HHS
  • The Search for a Field Sales Management Tool
  • How To Respond When a Contractor Goes Rogue: A Crisis Management Guide
  • Attorney Fees in Litigation, the Prevailing Party, Fee Applications, and Strategy (Part 2)
  • Working Together To Ensure No Kid Fights Cancer Alone

Also in TTUSA

  • A Jump in Used Car Prices for First Time in Six Months
  • Ten Proven Strategies to Boost Your Productivity
  • Catching Up with Alex Trebek and Friends in War on Pancreatic Cancer
  • How to Align Your Heart, Mind, Body, and Action with the Truth
  • Creating an Advertising Video Clip for Your Business From ABD-Video

RSS The Daily Blaze

  • Building Client Trust and the Human Connection
  • The Eaton and Palisades Fire One Year Later
  • Your Long Term Care Insurance Policy Doesn’t Cover Errands
  • Where Will Venezuela’s Dictator Flee to if US Demands Regime Change?
  • Trump’s ‘Media Bias’ Portal Uses Crowdsourcing for Fresh Propaganda

RSS USA Business Radio

  • Today’s Top Leadership Challenges—and How Leaders Navigate Them
  • The Ghost of Christmas Past and Enterprise Data
  • A Real World Look at Succession Law
  • Experiential Adventure Travel Through India and the Rest of South Asia!
  • The Absence of Good Faith, “Taylor’s Version” and the Pendulum

RSS USA Daily Times

  • International Bestselling Author on Her Latest Jewish Romance Novel
  • 5 Most Profitable Small Businesses in the UK for Fresh Graduates With Low Investment
  • Beyond Command: Lead With Flow & Momentum
  • Luxury Travel Within Reach
  • Veterans Day: A Time for Reflection and Responsibility

RSS USA Daily Chronicles.

  • Life After Ownership – Planning Your Purposeful Next Chapter
  • National Diabetes Month Spotlight
  • 10 Ethical ChatGPT Prompts for Answering Assignments Every Student Can Use (2025–26 Guide)
  • The Price of Pet Food
  • Part One: Rethinking Nutrition in America — a Conversation With Marion Nestle, Ph.D., M.P.H.

RSS Price of Business

RSS US Daily Review

  • Why Did Trump Reject Maduro’s Offer?
  • Unboxing Trump’s View of Affordability
  • A Unique Approach to Business Problem Solving
  • The Media Seems to Now Be Asking Tough Questions of Trump
  • Turbocharged Results: High-Impact Diesel Performance Components for Stronger Engines

PoB Digital Network

US Daily Review

USA Business Radio

USA Daily Chronicles

USA Daily Times

The Daily Blaze

The Times USA

Price of Business

Privacy Policy

https://www.thetimesusa.com/privacy-policy-2/

© 2025 The Times USA | Powered by Superbs Personal Blog theme