Showing posts with label data. Show all posts
Showing posts with label data. Show all posts

Monday, October 7, 2024

The Friend Network


Targeting friends to induce social contagion can benefit the world, says new research
May 2024, phys.org

The study evaluated a strategy that exploits the so-called "friendship paradox" of human social networks. That theory suggests that on average, your friends have more friends than you do. As the theory goes, the individuals nominated as friends potentially wield more social influence than those who identify them.

For the study, the researchers utilized the friendship paradox in the delivery of a proven 22-month education package promoting maternal, child, and neonatal health in 176 isolated villages in Honduras.

The researchers found that delivering the intervention to a smaller fraction of households in each village via the friendship targeting strategy led to the same level of behavioral adoption as would have been achieved by treating all the households.

People were either selected randomly within each village to receive the intervention or they were randomly chosen to nominate their friends, who were subsequently picked at random. 

"We found that targeting people's friends for an intervention induced significant social contagion, creating cascades of beneficial health practices to people who didn't receive the intervention." 

For many outcomes, using the friendship-nomination targeting method to reach 20% of households in a village affected outcomes the same as administering the intervention to every household.

Yes, Facebook knows this very well.

via Yale and Temple University: Edoardo M. Airoldi et al, Induction of social contagion for diverse outcomes in structured experiments in isolated villages, Science (2024). DOI: 10.1126/science.adi5147



Study shows relatively low number of superspreaders responsible for large portion of misinformation on Twitter
May 2024, phys.org

10 months of data; 2,397,388 tweets; 448,103 users; parsed by low-credibility information status.

A third of the low-credibility tweets had been posted by people using just 10 accounts, and just 1,000 accounts were responsible for posting approximately 70% of such tweets.

via Indiana University: Matthew R. DeVerna et al, Identifying and characterizing superspreaders of low-credibility content on Twitter, PLOS ONE (2024). DOI: 10.1371/journal.pone.0302201

Monday, May 15, 2023

Megadata vs Magadata


The death of open access mega-journals?
Mar 2023, phys.org

"...explosive growth of mega-journals may be accompanied by the fall of some previously prestigious journals."

Many newer mega-journals have begun specializing in discipline-focused journals that are publishing faster and in greater volume than traditional journals can keep up with.

And because getting more citations and publishing more stories in a current year helps lift the impact factor, self-citing journals are skewing the imapct factor of the journal. 

Using an internally-developed AI tool to help identify outlier characteristics that indicate that a journal may no longer meet quality criteria, the Web of Science has removed the impact factor of nearly two dozen journals, including one of the world's largest, the International Journal of Environmental Research and Public Health. Many of the journals published by Hindawi and two by MDPI have had their impact factor ratings removed, likely reflecting concerns with the integrity of the publishing process.
  • Hiring "guest editors" who may not be reviewing studies in their field of expertise
  • Quick turnaround times from submitting a paper to publication (200 hundred days in traditional publishing, 30 for Hindawi)
  • Hindawi was purchased by Wiley publishing in 2021 for $300 million and has already had to deal with thousands of retractions after uncovering thousands of fraudulent papers filled with off-subject citations.

via opinion letter by researchers from Italy and Stanford: John P. A. Ioannidis et al, The Rapid Growth of Mega-Journals Threats and Opportunities, JAMA (2023). DOI: 10.1001/jama.2023.3212

Megajournals may perpetuate and accentuate an already dysfunctional system of scientific evaluation and publication,” they write. The pay-for-publication model creates an incentive for authors trying to meet institutions’ quotas for publications, and “megajournals may drain an already strained pool of reviewers from traditional journals.” Ioannidis calls for more research comparing the quality of peer review in megajournals and traditional ones, and he suggests institutions and funders reward researchers for studies that are transparent and rigorous. 
-Fast-growing open-access journals stripped of coveted impact factors: Web of Science delists some 50 journals, including one of the world’s largest, Mar 2023, Jeffrey Brainard for Science [link]

And by the way:

AI language models open a potential Pandora's box of medical research fraud
Mar 2023, phys.org

Meta things -- they wanted to see if artificial intelligence could write a fabricated research paper and then investigate how best to detect it, so they ran their AI-generated text through a free, online, AI rephrasing tool -- the consensus unanimously flipped to "likely human," suggesting we need better AI detection tools, since these technologies could be used to write entire studies with false data, nonexistent participants and meaningless results.

via Faisal Elali of the State University of New York Downstate Health Sciences University: Faisal R. Elali et al, AI-generated research paper fabrication and plagiarism in the scientific community, Patterns (2023). DOI: 10.1016/j.patter.2023.100706


Thursday, May 4, 2023

Double Digital


DfAI: The missing piece of artificial intelligence engineering
Jan 2023, phys.org
https://techxplore.com/news/2023-01-dfai-piece-artificial-intelligence.html

They don't mention digital twins but they should:

DfAI - Design for Artificial Intelligence 

It's a data-rich process that captures intelligence throughout the lifecycle of the design, whether it's an airplane or an umbrella. But not enough companies are using it. This group says we need to raise AI literacy in industry, redesign engineering systems to better integrate with AI; and enhance the engineering AI development process.

And as robots take our jobs, they do tend to create new jobs in the process. Enter the "design repository curator".

via Carnegie Mellon University Mechanical Engineering and Re:Build Manufacturing: Glen Williams et al, Design for Artificial Intelligence: Proposing a Conceptual Framework Grounded in Data Wrangling, Journal of Computing and Information Science in Engineering (2022). DOI: 10.1115/1.4055854



How digital twins could protect manufacturers from cyberattacks
Feb 2023, phys.org

In addition to spotting routine indicators of wear and tear, digital twins could help find something more within manufacturing data, the authors of the study say.

"Because manufacturing processes produce such rich data sets -- temperature, voltage, current -- and they are so repetitive, there are opportunities to detect anomalies that stick out, including cyberattacks," 

via National Institute of Standards and Technology: E. C. Balta et al, Cyber-Attack Detection Digital Twins for Cyber-Physical Manufacturing Systems. IEEE Transactions on Automation Science and Engineering (2023). DOI: 10.1109/TASE.2023.3243147


Digital twin opens way to effective treatment of inflammatory diseases
Feb 2023, phys.org

In an inflammatory disease like rheumatoid arthritis, Crohn's disease or ulcerative colitis, thousands of genes alter the way they interact in different organs and cell types. Moreover, the pathological process varies from one patient to another with the same diagnosis, and even within the same patient at different times.

Every physiological process can be described with mathematical equations [earlier they call them "molecular programs"]. This advanced digital modeling technique can be adjusted to a patient's unique circumstances by analyzing the activity of each and every gene in thousands of individual cells from blood and tissue. Such a digital twin can be used to calculate the physiological outcome if a condition changes, such as the dosage of a drug.

In the current study, the researchers combined analyses of a mouse model of rheumatoid arthritis and digital twins of human patients with various inflammatory diseases.
"Even though only the joints were inflamed in mice, we found that thousands of genes changed their activity in different cell types in ten organs, including the skin, spleen, liver and lungs," says Dr. Benson. "As far as I'm aware, this is the first time science has obtained such a broad picture of how many organs are affected in rheumatoid arthritis. This is partly due to the difficulty of physically sampling so many different organs."

via Karolinska Institute Department of Clinical Science, Intervention and Technology: Mikael Benson, Multi-organ single cell analysis reveals an on/off switch system with potential for personalized treatment of immunological diseases, Cell Reports Medicine (2023). DOI: 10.1016/j.xcrm.2023.100956


Advances in brain modeling open a path to digital twin approaches for brain medicine
Mar 2023, phys.org

To create personalized brain models, the researchers use a simulation technology called The Virtual Brain (TVB), which HBP scientist Viktor Jirsa has developed together with collaborators. For each patient, the computational models are created from data of the individually measured anatomy, structural connectivity and brain dynamics.

via Human Brain Project and AMU Marseille: Viktor Jirsa et al, Personalised virtual brain models in epilepsy, The Lancet Neurology (2023). DOI: 10.1016/S1474-4422(23)00008-X

Wednesday, September 7, 2022

Weaponized Delivery Packages for Misinformation


One day customers will only want to do business with those who harvest their data sustainably.

Twitter pays $150M fine for using two-factor login details to target ads
May 2022, Ars Technica

"As the complaint notes, Twitter obtained data from users on the pretext of harnessing it for security purposes but then ended up also using the data to target users with ads," Federal Trade Commission Chair Lina Khan said. "This practice affected more than 140 million Twitter users, while boosting Twitter's primary source of revenue."


Feds seize SSNDOB marketplace that listed personal data of 24 million people
Jun 2022, Ars Technica

Social Security Number Date of Birth (SSNDOB) like the walmart of personal data.

More fallout from the Chainalysis revelation, which is basically that every transaction you make on the blockchain is public knowledge, so with some good network software, you can track people and money like a first grade math problem. 

Further readings:
Inside the Bitcoin Bust That Took Down the Web’s Biggest Child Abuse Site
Apr 2022, Mike McQuade, WIRED [soft paywall]


Facebook is receiving sensitive medical information from hospital websites
Jun 2022, The Markup via Ars Technica

Experts say some hospitals’ use of an ad tracking tool may violate a federal law protecting health information. (You don't say)

A tracking tool installed on many hospitals’ websites has been collecting patients’ sensitive health information — including details about their medical conditions, prescriptions, and doctor’s appointments — and sending it to Facebook.

The Markup tested the websites of Newsweek’s top 100 hospitals in America. On 33 of them we found the tracker, called the Meta Pixel, sending Facebook a packet of data whenever a person clicked a button to schedule a doctor’s appointment. 

Clicking the “Schedule Online” button, filling in the booking form, or clicking the “Finish Booking” button on a doctor’s page sent the following information:
  • text of the button clicked
  • doctor’s name
  • doctor's field of medicine
  • search term used to find doctor: “pregnancy termination"
  • condition selected from dropdown menu: “Alzheimer’s”
  • first name
  • last name
  • email address
  • phone number
  • zip code
  • city of residence  entered into the booking form,
  • names of patients’ medications
  • descriptions of their allergic reactions
  • upcoming doctor’s appointments
  • name and dosage of a medication in our health record
  • notes we had entered about the prescription
  • response to a question about sexual orientation
The Markup also found the Meta Pixel installed inside the password-protected patient portals of seven health systems. 
You heard the man; this is a stick up. 

Technical sidenote:
“The evil genius of Facebook’s system is they create this little piece of code [the pixel] that does the snooping for them and then they just put it out into the universe and Facebook can try to claim plausible deniability,” said Alan Butler, executive director of the Electronic Privacy Information Center. “The fact that this is out there in the wild on the websites of hospitals is evidence of how broken the rules are.” (So the pixel is like a dematerialized AirPod?)

Further reading on body brokers and biodata:
Sapiens For Sale, Network Address, Aug 2022

Network structure of Agents in Tsuchiyu Onsen, Tohoku University, 2022


Kochava faces legal action over sale of location data
Aug 2022, BBC News

The company, founded in 2011, says on its website that it "complies with all user data privacy and consent regulations".

And they do, because there aren't any.


Data privacy bill would give you more control over info collected about you
Aug 2022, The Conversation via Ars Technica

"Excludes deidentified data"
-American Data and Privacy Protection Act (Frank Pallone, hello NJ)

How hard is it to 'de-anonymize' cellphone data? 

Not hard:

We study fifteen months of human mobility data for one and a half million individuals and find that human mobility traces are highly unique. In fact, in a dataset where the location of an individual is specified hourly and with a spatial resolution equal to that given by the carrier's antennas, four spatio-temporal points are enough to uniquely identify 95% of the individuals.
-Unique in the Crowd: The privacy bounds of human mobility. Yves-Alexandre de Montjoye et al. Sci Rep 3, 1376 (2013). https://doi.org/10.1038/srep01376

Four datapoints. That's 2013 by the way.

Post Script:
Here's the thing: we've all watched the promise of tech and the internet curdle into (at best) invasive, advertisement-saturated, rent-seeking bullshit and/or (at worst) weaponized delivery packages for misinformation, bigotry, and occasional incitements to genocide and violence. I think we're all reaching our saturation limit for being monetized, marketed to, invasively tracked, and charged a premium for devices and services that enable those things. I think we're looking for relief from all that, not variety of opportunities to experience it. -Snark128, "Meta sparks anger by charging for VR apps", Financial Times via Ars Technica, Jun 2022 https://www.ft.com/content/e8910bad-b873-407d-b1ca-46eb4ceb3db2 


Monday, March 28, 2022

Watch Words Come To Life


Turning an analysis of Asimov's Foundation into art
Oct 2021, phys.org

Just a great example of cross-disciplinary data art and network science, via Datapolis and the Central European University in Budapest: Milán Janosov, Flóra Borsi, Asimov's foundation—turning a data story into an NFT artwork. arXiv:2109.15079v1 [physics.soc-ph], arxiv.org/abs/2109.15079

See also this thing which I don't understand called an NFT platform: https://foundation.app/@milanjanosov/~/92747

Some interesting bits noted in the article:

  • We found that among the most mentioned keywords there were three different planets.
  • Different planets play different roles in the book. That's why we started to look at the emotional arcs of these planets.
  • We extracted a series of sentences about each planet, and we used a happiness scoring algorithm to create their emotional arc.
  • The arc of Trantor was going down, and that coincided with the fact that both Trantor and the galactic were also falling during the series; while on the other hand, Terminus, which is the heart of the Foundation, slowly rises.

Then they take an 8,000-word network graph and run it in the chronological order of the book, with the planets linking the words associated with them, but rising and falling throughout the arc of the narrative. So now it's a video. And they made their own soundtrack.

Post Script:
How does the brain interpret computer languages?
Mar 2021, Ars Technica

Interestingly, code-solving activated parts of the multiple-demand network that are not activated when solving math problems. So the brain doesn’t tackle it as language or logic -- it appears to be its own thing.

via MIT, Tufts: Comprehension of computer code relies primarily on domain-general executive brain regions. Anna A Ivanova et al. eLife, 2020. DOI: 10.7554/eLife.58906 


Post Post Script:
In a neuroprosthetic first, ALS patient sends social media message via brain-computer interface
Jan 2022, phys.org

It's called the Stentrode Brain Computer Interface, developed by brain computer interface company Synchron.

The real news here is that it's blurred the line over what we call invasive neural implants, since it was snaked through his jugular vein. This allows us to get really good signals from your brain without having to drill a hole in your head. 

via Synchron Press Release, "First Tweet by Implanted BCI", Dec 2021



Monday, March 14, 2022

Look Mom No Data


AKA From Deep Learning to Deep Reasoning

DRNets can solve Sudoku, speed scientific discovery
Sep 2021, phys.org

You can teach a machine to recognize a dog by showing it 1,000 pictures of dogs, Gomes said, but scientific discovery is not like that.

"You are not going to have lots and lots of labeled data," she said. "And in general, the examples you have are not exactly what you are looking for, but then you reason about what you know scientifically about the domain, and you can infer new knowledge."

Key to DRNets is the idea of an "interpretable latent space." Basically, it gives DRNets the ability to reason about the constraints of the domain—in this case materials science—from input data.

They started with Sudoku -- de-mixing overlapping handwritten Sudoku puzzles—grids. The computer had to separate the puzzles into two solved Sudokus, without any training data, which it was able to achieve with close to 100% accuracy.

The researchers then put DRNets to work on a real-world problem: automating crystal-structure phase mapping of solar-fuels materials, using X-ray diffraction (XRD) patterns. Crystal-structure phase mapping involves separating the source XRD signals of the desired crystal structures from "noisy" mixtures of XRD patterns, a task for which labeled training data are typically not available. ... DRNets was able to identify and separate a total of 13 crystal phases (single-phase materials) in 19 unique mixtures of the single-phase materials. ... DRNets' findings, verified using manual analysis, enable the discovery of complex mixtures of crystalline materials that convert solar energy into storable solar chemical fuels.

via Cornell University: Di Chen et al, Automating crystal-structure phase mapping by combining deep learning with constraint reasoning, Nature Machine Intelligence (2021). DOI: 10.1038/s42256-021-00384-1


Friday, August 18, 2017

Unzipf


Zipf's law, top ten most favorite thing on Network Address. New theory ---

Unzipping Zipf's Law: Solution to a century-old linguistic problem
Aug 2017, phys.org

Sander Lestrade, a linguist at Radboud University in The Netherlands, proposes a new solution to this notorious problem in PLOS ONE.

...shows that Zipf's law can be explained by the interaction between the structure of sentences (syntax) and the meaning of words (semantics) in a text.

"In the English language, but also in Dutch, there are only three articles, and tens of thousands of nouns," Lestrade explains. "Since you use an article before almost every noun, articles occur way more often than nouns." But that is not enough to explain Zipf's law. "Within the nouns, you also find big differences. The word 'thing', for example, is much more common than 'submarine', and thus can be used more frequently. But in order to actually occur frequently, a word should not be too general either. If you multiply the differences in meaning within word classes, with the need for every word class, you find a magnificent Zipfian distribution. And this distribution only differs a little from the Zipfian ideal, just like natural language does.
-phys.org

WHAT'S ZIPF

The most frequent word in a language, or in a book, or whatever, will occur approximately twice as often as the second most frequent word, three times as often as the third most frequent word, etc.

(straight from wikipedia, I mean it's all numbers anyway, right?)

For example, in the The Brown University Standard Corpus of Present-Day American English, the word "the" is the most frequently occurring word, and by itself accounts for nearly 7% of all word occurrences (69,971 out of slightly over 1 million). True to Zipf's Law, the second-place word "of" accounts for slightly over 3.5% of words (36,411 occurrences), followed by "and" (28,852). Only 135 vocabulary items are needed to account for half the Brown Corpus.

The same relationship occurs in many other rankings unrelated to language, such as the population ranks of cities in various countries, corporation sizes, income rankings, and so on.
http://en.wikipedia.org/wiki/Zipf's_law

*Zipf's law is referenced in Science Fiction author Robert J. Sawyer's www.wake, when the main character is searching for intelligent life on the web.
http://en.wikipedia.org/wiki/Wake_(Robert_J._Sawyer_novel)

META

There's some other laws meta-physical, like Benford's Law:

In this distribution, the number 1 occurs as the first digit about 30% of the time, while larger numbers occur in that position less frequently, with larger numbers occurring less often: 9 as the first digit less than 5% of the time. This distribution of first digits is the same as the widths of gridlines on a logarithmic scale.


POST SCRIPT
other meta-phys laws etc.

Bursts
Network Address, 2012

Laws Meta-Physical
Network Address, 2013

Physicists eye neural fly data, find formula for Zipf's law
August 2014, phys.org

mathematical models, which demonstrate how Zipf's law naturally arises when a sufficient number of units react to a hidden variable in a system.

"If a system has some hidden variable, and many units, such as 40 or 50 neurons, are adapted and responding to the variable, then Zipf's law will kick in."

"We showed mathematically that the system becomes Zipfian when you're recording the activity of many units, such as neurons, and all of the units are responding to the same variable".

Ilya Nemenman, biophysicist at Emory University and co-author
-phys.org

Monday, January 19, 2015

Magnonic Holographic Memory Device


Your Mind is Belong to Us

Researchers demonstrate holographic memory device
phys.org, Feb 2014


and on that note:

Scientists develop thought-controlled gene switch
BBC News - Nov 2014

Monday, July 1, 2013

Anthropogenic Metadata on Climate Science

aka The Low-Hanging Fruit of Neuro-Pop in the Age of Big Datty


Notice below, just a sample of the kinds of reports that use climate science as a substrate upon which to study human behavior and cognition..

Some may find it interesting to see how the science of climate change has ripened into a metadata-rich fruit hanging low on the tree of human-knowledge; that's knowledge -of- humans, not knowledge acquired by humans.

Granted we are entering the age of the New Humanities, and granted climate science is probably the most contentious scientific issue to be argued publicly since the advent of digital social media, but, 10 years ago, had you asked anyone what would be the most significant by-product of climate science, this would have never made it to the list.

Just goes to show that the world is full of surprises, and no matter how much you try, predicting the value of scientific endeavour is never part of the justification of doing it.

[note: I've recently been watching panels discussing the value of space exploration, and exploration in general]

Changing minds about climate change policy can be done—sometimes
phys.org Jun 24, 2013
Provided by The Ohio State University

Anthropologists argue field must play a vital role in climate change studies
phys.org Jun 20, 2013
Provided by University of California - Santa Cruz

Some Americans are cooling off on global warming
phys.org Jun 28, 2013
Provided by University of Michigan

Emotional response to climate change influences whether we seek or avoid further information
phys.org May 15, 2013
Provided by University at Buffalo

Americans care deeply about 'global warming' – but not 'climate change'
The Guardian, 27 May 2014

Sunday, November 18, 2012

A Reason for Irony


A response to The Stone article: How to Live Without Irony

credit: Leif Parsons

IRONY: “It pre-emptively acknowledges its own failure to accomplish anything meaningful. No attack can be set against it, as it has already conquered itself [do I even need to cite Anonymous here]. The ironic frame functions as a shield against criticism […] to dodge responsibility…to secretly flee…”
-OR, are we just unconsciously recognizing, finally, the fallacy of causality?

Irony is a defense, “a shield against criticism” but it also a weapon, or rather, a tool. It is what we use to sift through the bullshit. It is what we use to shift the power structures.

Value? From whence, for whom? In the old clothesline paradox, as value recedes, retreats, packs it bags and gets on a plane to Iceland, we use irony to make sure it doesn’t come back home. Back home to the banks and the corporations.

Irony is a shield, not just for us the users, but for the value itself, the lambswool that hides the wolf, but, in this case, not to penetrate and attack, instead to escape, like Iranian-crisis hostages.

It is not us who are hiding. We are hiding something, a game of hot-potato, or hide-and-seek, or just keep-away, just long enough for Big Everything and Big Everyone to lose track, or lose interest.

MEANING: “Moving away from the ironic involves saying what you mean, meaning what you say and considering seriousness and forthrightness as expressive possibilities, despite the inherent risks.” Sounds as if being absurd were a cop-out, yet the author says just prior, that being a hipster entails processing “several stages of self-scrutiny”.

In a world where it is almost possible to predict the future, (Barabasi’s Bursts, Sandy and the success of meteorological prediction, election forecasting, speculative futures, financial engineering, predictive analysis, etc…), how else can you avoid being a foregone conclusion unless you completely make no sense.

Absurdity is the antidote to the probabilistic world. Interestingly, Hipsters ^here are referred to as Harlequins; John Twelve Hawks, in his 2005 future-fiction, also writes of  Harlequins’ using random number generators to help them make decisions randomly, thus subverting the Vast Machine (or as PKD called it, the Vast Active Living Information System).

If Big Data has given us anything thus far, it is “The Search for Meaning”, and if culture has offered any comfort, any insulation, as is its purpose, it is giving us new ways of meaning, new ways of being.

How do you make meaning out of absurdity? That, of course, is the new frontier. Artists, in all regards, are  always one step ahead; Hipsters, with their slieght-of-hand tautologies, giving us a glimpse of the future, ghettouflaged in junk data – an encrypted program for a ‘way of living’: a generation not-yet has the cipher.


partially unrelated image

MODUS OPERANDI:
As my own page for a few years now has read in its subtitle: ‘Clarification, Contradiction, and Confusion’, I somehow seemed to have fallen victim to this pre-emptive defense. Or have I?

I gave feliz navidad christmas cards to my family for years (we don’t speak Spanish). It is not because I fear dislike, however, it is because I’ve tried to give away all my fucks, like an anti-Scrooge McDuck.

If I do not speak in the language of the world around me, how can I live? And yet, as such, I write, for cultural critics of the future to point out with effortless accuracy, the Greta Garbo Lips of my portraits.
NOTES:
How to Live Without Irony
CHRISTY WAMPOLE, THE STONE November 17, 2012

Bursts: Can Human Behavior Be Predicted And Controlled?
Albert-Laszlo Barabasi, 2010

“Sandy shows storm-prediction progress”, Business of Federal Technology
Frank Konkel, The, Nov 05, 2012

[Argo, the 2012 film]


Stevens Institute, “Financial Engineering”

The Traveler, John Twelve Hawks, 2005

VALIS, Philip K. Dick, 1981

Greta Garbo Lips and the Van Meegeren effect, in:
The Forger’s Art, Denis Dutton, ed., 1983
http://books.google.com/books/about/The_Forger_s_Art.html?id=wV_CHAAACAAJ

POST SCRIPT:

U.S. Cities Relying on Precog Software to Predict Murder
KIM ZETTER 01.10.13
The software parses about two dozen variables, including criminal record and geographic location. The type of crime and the age at which it was committed, however, turned out to be two of the most predictive variables.
“People assume that if someone murdered then they will murder in the future,” Berk told the news outlet. “But what really matters is what that person did as a young individual. If they committed armed robbery at age 14 that’s a good predictor. If they committed the same crime at age 30, that doesn’t predict very much.”
-Richard Berk, criminologist at the University of Pennsylvania who developed the algorithm
http://www.wired.com/threatlevel/2013/01/precog-software-predicts-crime/