We use cookies, including third-party cookies from Google to serve personalized ads through AdSense, to operate this site and understand how it is used. By continuing to browse, you accept this use. See our Privacy Policy and Terms of Use for details, including how to opt out of personalized advertising.
Accept
SmartData CollectiveSmartData Collective
  • Analytics
    AnalyticsShow More
    The 10 Best Analytics Tools in 2026, Grouped by Job -- AI-generated illustration
    The 10 Best Analytics Tools in 2026, Grouped by Job
    36 Min Read
    Data Modeling Tools: 14 Picks Compared by Modeling Layer in 2026 -- AI-generated illustration
    Data Modeling Tools: 14 Picks Compared by Modeling Layer in 2026
    42 Min Read
    chatgpt image jul 21, 2026, 04 34 30 pm
    4 Core Benefits of Predictive Maintenance after Vibration Analysis
    10 Min Read
    How Does Data Mining Boost Customer Satisfaction in Logistics? Harnessing Analytics for Results -- AI-generated illustration
    How Does Data Mining Boost Customer Satisfaction in Logistics? Harnessing Analytics for Results
    11 Min Read
    chatgpt image jul 13, 2026, 04 23 45 pm
    How Data Analytics Helps Companies Improve User Engagement
    19 Min Read
  • Big Data
  • BI
  • Exclusive
  • IT
  • Marketing
  • Software
Search
© 2008-25 SmartData Collective. All Rights Reserved.
Reading: The Unreasonable Effectiveness of Data
Share
Notification
Font ResizerAa
SmartData CollectiveSmartData Collective
Font ResizerAa
Search
  • About
  • Help
  • Privacy
Follow US
© 2008-23 SmartData Collective. All Rights Reserved.
SmartData Collective > Uncategorized > The Unreasonable Effectiveness of Data
Uncategorized

The Unreasonable Effectiveness of Data

Daniel Tunkelang
Daniel Tunkelang
4 Min Read
SHARE

Over the past week, there’s been lots of commentary about “The Unreasonable Effectiveness of Data“, an article by Googlers Alon Halevy, Peter Norvig, and Fernando Pereira in the most recent issue of IEEE Intelligent Systems.

Here are a few posts that have been appearing in my RSS reader:

  • Geeking with Greg: Semantic interpretation and the effectiveness of big data
  • Jeff’s Search Engine Caffe: Statistical Learning of Semantics from Web Data
  • Matthew Hurst: Strings are not Meanings
  • Stefano’s Linotype: Unreasonable Hypocrisy

I’m intrigued by the amount of attention this paper has attracted–especially the vitriol in this Stefano’s post:

More Read

3 Ways your Technology Should Service Customer Experience in 2016
Apache ODE is Reaching Critical Mass
IT Project Success: There’s Magic In the Middle
How Enterprise 2.0 Will Enable the Semantic Web
A Sound Approach to Exploratory Music Search?

What upset me about that paper is not how they say “oh sure, structure is great, but look overhere: there is a goldmine in all the sand” (which is something I fully resonate with) but they phrased it as a fight, deterministic vs. statistical, trying to convince people that adding structure it not the way to go, it’s basically a global waste of research resources.

And yet, without the <a> tag (that is: machine-readable imposed structure), they wouldn’t be where they are, not they would be able to speak from such a tall soapbox.

I’m actually sympathetic to the view that it’s usually better to have more data than heavier theoretical machinery. But I’ve seen this view taken to an extreme so absurd as to be worthy of an April Fool’s joke–in Chris Anderson’s Wired article about “The End of Theory“. Moreover, that same article quotes Peter Norvig as saying that “All models are wrong, and increasingly you can succeed without them.”

So perhaps Stefano is right to react so harshly.

Link to original post

Share This Article
Facebook Pinterest LinkedIn
Share

Follow us on Facebook

Latest News

The 10 Best Analytics Tools in 2026, Grouped by Job -- AI-generated illustration
The 10 Best Analytics Tools in 2026, Grouped by Job
Analytics Big Data Exclusive
Data Modeling Tools: 14 Picks Compared by Modeling Layer in 2026 -- AI-generated illustration
Data Modeling Tools: 14 Picks Compared by Modeling Layer in 2026
Modeling
5 Common Mistakes Businesses Make During the Risk Assessment Process -- AI-generated illustration
5 Common Mistakes Businesses Make During the Risk Assessment Process
Business Intelligence Exclusive Risk Management
Best Age Estimation Software in 2026: Which Facial Age Providers Actually Hold Up -- AI-generated illustration
Best Age Estimation Software in 2026: Which Facial Age Providers Actually Hold Up
Artificial Intelligence Exclusive Machine Learning

Stay Connected

1.2KFollowersLike
33.7KFollowersFollow
222FollowersPin

You Might also Like

‘Business-IT alignment’ is dead… whatever it was

1 Min Read

Celebrating New Year (Already)

4 Min Read

SOA may be taking the risk out of cloud computing

1 Min Read

Change.gov

2 Min Read

SmartData Collective is one of the largest & trusted community covering technical content about Big Data, BI, Cloud, Analytics, Artificial Intelligence, IoT & more.

AI chatbots
AI Chatbots Can Help Retailers Convert Live Broadcast Viewers into Sales!
Chatbots
giveaway chatbots
How To Get An Award Winning Giveaway Bot
Big Data Chatbots Exclusive

Quick Link

  • About
  • Contact
  • Privacy
Follow US
© 2008-26 SmartData Collective. All Rights Reserved.
Welcome Back!

Sign in to your account

Username or Email Address
Password

Lost your password?