Intelligent Enterprise

Better Insight for Business Decisions

Intelligent Enterprise - Better Insight for Business Decisions
search Intelligent Enterprise
Advanced Search
RSS
Webcasts
Digital Library
Subscribe
Home



Hot Topics in High-Performance Analytics | Intelligent Enterprise Blog
Data Frontiers, by Curt Monash
Curt Monash runs Monash Research, which provides strategic, analysis-based advice to users and vendors of advanced information technology. He also writes the blogs DBMS2, Text Technologies, and Strategic Messaging.
See More by Curt Monash

E-MAIL  |  
Share
Hot Topics in High-Performance Analytics

Posted by Curt Monash
Monday, November 17, 2008
10:03 AM

For the past few months, I've collected a lot of data points to the effect that high-performance analytics – i.e., beyond straightforward query — is becoming increasingly important. And I've written about some of them at length. For example:

Ack. I can't decide whether "analytics" should be a singular or plural noun. Thoughts?

Another area that's come up which I haven't blogged about so much is data mining in the database. Data mining accounts for a large part of data warehouse use. The traditional way to do data mining is to extract data from the database and dump it into SAS. But there are problems with this scenario, including:

  • There's a lot of data to move.
  • Therefore it's tempting to only sample the database rather than analyze the whole thing, which could have at least a slight negative effect on model accuracy.
  • The result of the process is often some kind of scoring algorithm, and you may want to execute that real-time rather than in batch mode.

Various interesting fixes have been tried.

  • SAS and Teradata are partnering quite closely to run SAS on Teradata boxes.
  • Database management system vendors are building at least the data scoring part right into the DBMS. SAS rival SPSS – which relies more on just-in-time SQL and less on batch extracts anyway – reports that hooking into Oracle's native scoring produces massive performance gains. (To put that another way – I finally got independent confirmation of what Oracle's Charlie Berger has been telling me for years.)
  • Data preparation can be handled by the general ELT/ETLT (Extract/(Transform)/Load/Transform – i.e., in-database data transformation) strategies of the data warehouse DBMS vendors.
  • Oracle (more than most competitors, although SAS/Teradata are headed that way too) actually does all stages of data mining right in the database.

Vendors who are putting considerable marketing emphasis on parallel analytics include:

I'm sure others would say they belong on the list as well. It's an important area of competitive differentiation.



E-MAIL  |  
Share




This is a public forum. United Business Media and its affiliates are not responsible for and do not control what is posted herein. United Business Media makes no warranties or guarantees concerning any advice dispensed by its staff members or readers.

Community standards in this comment area do not permit hate language, excessive profanity, or other patently offensive language. Please be aware that all information posted to this comment area becomes the property of United Business Media LLC and may be edited and republished in print or electronic format as outlined in United Business Media's Terms of Service.

Important Note: This comment area is NOT intended for commercial messages or solicitations of business.


 




    Subscribe to RSS feed of all blogs


 



InformationWeek Business Technology Network
InformationWeekInformationWeek 500InformationWeek 500 ConferenceInformationWeek AnalyticsInformationWeek CIO
InformationWeek EventsInformationWeek ReportsInformationWeek MagazinebMightyByte and SwitchDark Reading
Digital LibraryIntelligent EnterpriseInternet EvolutionNetwork ComputingNo JitterPlug Into The Cloud
space
Techweb Events Network
InteropVoiceConWeb 2.0 ExpoWeb 2.0 SummitEnterprise 2.0 ConferenceMobile Business ExpoSoftware ConferenceCSI - Computer Security Institute
Black HatGTECEnergy CampMashup CampStartup Camp
space
Light Reading Communications Network
Light ReadingLight Reading EuropeUnstrungLight Reading's Cable Digital NewsConstantinopleInternet EvolutionPyramid Research
Heavy ReadingLight Reading Live!Light Reading InsiderEthernet ExpoOptical ExpoTeleco TVTower Technology Summit
space
Financial Technology Network
Advanced TradingBank Systems & TechnologyInsurance & TechnologyWall Street & TechnologyAccelerating Wall StreetBank Systems & Technology Executive SummitBuyside Trading SummitInsurance & Technology Executive Summit
space
Microsoft Technology Network
MSDN MagazineTechNetThe Architecture Journal
space