Welcome!

Machine Learning Authors: Elizabeth White, Yeshim Deniz, Liz McMillan, Jyoti Bansal, Kevin Benedict

Related Topics: Agile Computing, Industrial IoT, Open Source Cloud, Cognitive Computing , Machine Learning , Cloud Security

Agile Computing: Blog Post

Using Taxonomy to Drive Online Contextual Advertising with Sophializer

Classifying Web Content to the IAB Taxonomy

It’s a Big Market …
…  online advertising.  There are 10,000 stories and data points about it.  Here are two to give some context to the journey below.  First, global online ad spending is projected by ZenithOptimedia to exceed print ad spend by 2015 (note 1).  This 2015 projected spend figure for online advertising is $132.4 billion.  Second, global online ad revenue is projected by another research agency, Digital TV Research, to hit $143 billion by 2017 (note 2).

These are prodigious amounts of money for companies to spend to connect with customers.  But … surely it’s easy to connect online customers to web content featuring, or suggesting, products? And surely, online is “better”?  Where can, and do, taxonomy-based approaches add value to this dance of moving (emotional and semantic) parts between the intentful consumer poised to shop and the intentful marketer with honed content?

Online Ad Targeting is Easy … so 'They' Say …
Really?  So what might be “easy”?  And, indeed, “better”?  Let’s unbundle these simulacra that look like very fuzzy concepts, and as ontologists and knowledge engineers let’s think our way forward with the concept of “precision”.

So … online is more precise than billboards by freeways?  Lightly stated, online has advantages.  What about magazine print ads vs. online?  Online has potential advantages. But … and this is a very big but … in both these cases (and all others) online depends on connecting potential customers to products, their features, their benefits, their attributes and so on precisely, and with precision that is repeatable and extensible.  Rather than random (random is the most expensive way to advertise and has fallen out of favor).  And, since online copy and online ads are words (including in videos) and are semantically classifiable, and since classifications can be organized into models (taxonomies and ontologies) … then there are advantages to be created through the combination of semantic analysis, categorization and taxonomy.

Now, let’s connect taxonomy, classification, semantics and optimizing online ad targeting.  There are a host of holy grails currently being sought in the web/mobile/social uber-ecosystem.  Some are well found, though not perfect, and are unlikely to traverse through a paradigmatic improvement.  Think ‘search’.  Others are most definitely not found (yet).  Given the size of the market outlined in the first paragraph, the rewards are huge to those with the tools and skillsets that know how to work with semantics, taxonomy/ontology, classification of content to taxonomy, and design of taxonomies to drive online targeting.

New Approaches to Classifying to the IAB Taxonomy with Sophializer
Sophia Search
is a recent entrant into this space.  (I have written about them before here.   Sophia Search’s tool – currently called the ‘Sophializer’ – categorizes any URL to nodes in the Internet Advertising Bureau (IAB) taxonomy.  Sophializer can also classify content of ads (and so create a semantic/conceptual ‘signature’ for each). The IAB Contextual Taxonomy comprises three levels:

  • Tier 1 – 23 nodes
  • Tier 2 – 371 nodes
  • Tier 3 – unspecified and vendor specific

Given that Sophializer categorizes both sides of this content dance – web page and ad – web properties can serve ads to any page automatically using the IAB taxonomy as the cross-mapping conceptual foundation.

Sophializer not only classifies to Tier 1 and Tier 2 it also discovers/generates robust classifications that can be used to customize Tier 3 for individual customers.

Benefits of Using Taxonomy for Ad Targeting
Taxonomy gives a framework to this kind of semantic work.  Essentially, we are cross-mapping both partners of this content dance – content and ad - using the IAB taxonomy  as a “choreographer” of sorts.  Other taxonomies could be used.  In fact, multiple taxonomies could be used – and this would be particularly powerful if these taxonomies were cross-mapped to each other.  For example, if you have content (web page, say, or ad) categorized and mapped to Taxonomy A and Taxonomy A is cross-mapped to the IAB taxonomy … then … you can propagate these ads to content that is already categorized.

Benefits of Using Categorization Tools to Assign Marketing Content to Taxonomy Nodes
There are a number of different methods of assigning content to nodes in any taxonomy –

  • Manually
  • Training sets of documents (training documents are most often manually selected as exemplars)
  • Categorization algorithms that work with semantic tokens

There is more than enough to say on each of these around methods, workflows, best practices and pitfalls for a blog post on each.  But not here.

Sophializer utilizes patented and proprietary algorithms in the core of their categorization engine.  Two fundamental points are worth, briefly, focusing on.  Firstly, different categorization engines use different patented technologies.  “Quality” from different categorizers is (very) variable.  Which is why it is important to carry out “Proofs of Concept” when evaluating this technology.

Secondly, the more semantically rich the taxonomy – e.g. fully enriched with synonyms and other types of evidence terms – the better “quality” one gets with any method of associating content to taxonomy nodes.   Both of these parameters are make-or-break (literally) in using semantics to target online ads.

Learn More 2.0
The Google Display Network is IAB Certified and complies with the top 2 tiers of the IAB Contextual Taxonomy.  You can read details of what Google do here and this also navigates you to the Google mapping to the IAB taxonomy Tier 1 and Tier 2.

Sophia Search currently has a number of engagements on the web that are live.  For example, targeting ads for non-fiction books (from a major publishing house) to news stories (on a pre-eminent news site).  You can contact them for details.

This is not an empty space.  Other companies are also searching for the holy grail of taxonomy-based content targeting mediated by content categorization that works.  See, for example, see ADmantX (http://blog.admantx.com/post/15726823528/a-new-iab-based-taxonomy-and-an...).

This whole space is an excellent example of where the application of the nexus of taxonomy, categorization and semantics will provide stratospheric business benefit.  Grails are waiting to be found here.

Notes
Note 1.  See ZenithOprimedia

The detailed ZenithOptimedia figures can be found here

Note 2.  See Hollywood Reporter

You can download the Digital TV Research press release about these figures here

@CloudExpo Stories
Apache Hadoop is emerging as a distributed platform for handling large and fast incoming streams of data. Predictive maintenance, supply chain optimization, and Internet-of-Things analysis are examples where Hadoop provides the scalable storage, processing, and analytics platform to gain meaningful insights from granular data that is typically only valuable from a large-scale, aggregate view. One architecture useful for capturing and analyzing streaming data is the Lambda Architecture, represent...
SYS-CON Events announced today that Ocean9will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Ocean9 provides cloud services for Backup, Disaster Recovery (DRaaS) and instant Innovation, and redefines enterprise infrastructure with its cloud native subscription offerings for mission critical SAP workloads.
With billions of sensors deployed worldwide, the amount of machine-generated data will soon exceed what our networks can handle. But consumers and businesses will expect seamless experiences and real-time responsiveness. What does this mean for IoT devices and the infrastructure that supports them? More of the data will need to be handled at - or closer to - the devices themselves.
In his session at DevOps Summit, Tapabrata Pal, Director of Enterprise Architecture at Capital One, will tell a story about how Capital One has embraced Agile and DevOps Security practices across the Enterprise – driven by Enterprise Architecture; bringing in Development, Operations and Information Security organizations together. Capital Ones DevOpsSec practice is based upon three "pillars" – Shift-Left, Automate Everything, Dashboard Everything. Within about three years, from 100% waterfall, C...
Adding public cloud resources to an existing application can be a daunting process. The tools that you currently use to manage the software and hardware outside the cloud aren’t always the best tools to efficiently grow into the cloud. All of the major configuration management tools have cloud orchestration plugins that can be leveraged, but there are also cloud-native tools that can dramatically improve the efficiency of managing your application lifecycle.
SYS-CON Events announced today that Juniper Networks (NYSE: JNPR), an industry leader in automated, scalable and secure networks, will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Juniper Networks challenges the status quo with products, solutions and services that transform the economics of networking. The company co-innovates with customers and partners to deliver automated, scalable and secure network...
DevOps is often described as a combination of technology and culture. Without both, DevOps isn't complete. However, applying the culture to outdated technology is a recipe for disaster; as response times grow and connections between teams are delayed by technology, the culture will die. A Nutanix Enterprise Cloud has many benefits that provide the needed base for a true DevOps paradigm. In his Day 3 Keynote at 20th Cloud Expo, Chris Brown, a Solutions Marketing Manager at Nutanix, will explore t...
SYS-CON Events announced today that SoftLayer, an IBM Company, has been named “Gold Sponsor” of SYS-CON's 18th Cloud Expo, which will take place on June 7-9, 2016, at the Javits Center in New York, New York. SoftLayer, an IBM Company, provides cloud infrastructure as a service from a growing number of data centers and network points of presence around the world. SoftLayer’s customers range from Web startups to global enterprises.
Some people worry that OpenStack is more flash then substance; however, for many customers this could not be farther from the truth. No other technology equalizes the playing field between vendors while giving your internal teams better access than ever to infrastructure when they need it. In his session at 20th Cloud Expo, Chris Brown, a Solutions Marketing Manager at Nutanix, will talk through some real-world OpenStack deployments and look into the ways this can benefit customers of all sizes....
Deep learning has been very successful in social sciences and specially areas where there is a lot of data. Trading is another field that can be viewed as social science with a lot of data. With the advent of Deep Learning and Big Data technologies for efficient computation, we are finally able to use the same methods in investment management as we would in face recognition or in making chat-bots. In his session at 20th Cloud Expo, Gaurav Chakravorty, co-founder and Head of Strategy Development ...
SYS-CON Events announced today that Linux Academy, the foremost online Linux and cloud training platform and community, will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Linux Academy was founded on the belief that providing high-quality, in-depth training should be available at an affordable price. Industry leaders in quality training, provided services, and student certification passes, its goal is to c...
SYS-CON Events announced today that CA Technologies has been named “Platinum Sponsor” of SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY, and the 21st International Cloud Expo®, which will take place October 31-November 2, 2017, at the Santa Clara Convention Center in Santa Clara, CA. CA Technologies helps customers succeed in a future where every business – from apparel to energy – is being rewritten by software. From ...
Interoute has announced the integration of its Global Cloud Infrastructure platform with Rancher Labs’ container management platform, Rancher. This approach enables enterprises to accelerate their digital transformation and infrastructure investments. Matthew Finnie, Interoute CTO commented “Enterprises developing and building apps in the cloud and those on a path to Digital Transformation need Digital ICT Infrastructure that allows them to build, test and deploy faster than ever before. The int...
What if you could build a web application that could support true web-scale traffic without having to ever provision or manage a single server? Sounds magical, and it is! In his session at 20th Cloud Expo, Chris Munns, Senior Developer Advocate for Serverless Applications at Amazon Web Services, will show how to build a serverless website that scales automatically using services like AWS Lambda, Amazon API Gateway, and Amazon S3. We will review several frameworks that can help you build serverle...
SYS-CON Events announced today that Loom Systems will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Founded in 2015, Loom Systems delivers an advanced AI solution to predict and prevent problems in the digital business. Loom stands alone in the industry as an AI analysis platform requiring no prior math knowledge from operators, leveraging the existing staff to succeed in the digital era. With offices in S...
SYS-CON Events announced today that Interoute, owner-operator of one of Europe's largest networks and a global cloud services platform, has been named “Bronze Sponsor” of SYS-CON's 20th Cloud Expo, which will take place on June 6-8, 2017 at the Javits Center in New York, New York. Interoute is the owner-operator of one of Europe's largest networks and a global cloud services platform which encompasses 12 data centers, 14 virtual data centers and 31 colocation centers, with connections to 195 add...
The Software Defined Data Center (SDDC), which enables organizations to seamlessly run in a hybrid cloud model (public + private cloud), is here to stay. IDC estimates that the software-defined networking market will be valued at $3.7 billion by 2016. Security is a key component and benefit of the SDDC, and offers an opportunity to build security 'from the ground up' and weave it into the environment from day one. In his session at 16th Cloud Expo, Reuven Harrison, CTO and Co-Founder of Tufin, ...
SYS-CON Events announced today that T-Mobile will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. As America's Un-carrier, T-Mobile US, Inc., is redefining the way consumers and businesses buy wireless services through leading product and service innovation. The Company's advanced nationwide 4G LTE network delivers outstanding wireless experiences to 67.4 million customers who are unwilling to compromise on ...
SYS-CON Events announced today that HTBase will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. HTBase (Gartner 2016 Cool Vendor) delivers a Composable IT infrastructure solution architected for agility and increased efficiency. It turns compute, storage, and fabric into fluid pools of resources that are easily composed and re-composed to meet each application’s needs. With HTBase, companies can quickly prov...
SYS-CON Events announced today that Infranics will exhibit at SYS-CON's 20th International Cloud Expo®, which will take place on June 6-8, 2017, at the Javits Center in New York City, NY. Since 2000, Infranics has developed SysMaster Suite, which is required for the stable and efficient management of ICT infrastructure. The ICT management solution developed and provided by Infranics continues to add intelligence to the ICT infrastructure through the IMC (Infra Management Cycle) based on mathemat...